Minimal, Local, Causal Explanations for Jailbreak Success in Large Language Models
arXiv:2605.00123v1 Announce Type: new Abstract: Safety trained large language models LLMs can often be induced to answer harmful requests through jailbreak prompts.
AI 업계의 최신 소식을 빠르게 확인하세요.
arXiv:2605.00123v1 Announce Type: new Abstract: Safety trained large language models LLMs can often be induced to answer harmful requests through jailbreak prompts.
arXiv:2605.00136v1 Announce Type: new Abstract: Toolaugmented reasoning has become a popular direction for LLMbased agents, and it is widely assumed to improve reasoning and reliability.
arXiv:2605.00224v1 Announce Type: new Abstract: Aligning large language models LLMs with human preferences is commonly done via reinforcement learning from human feedback RLHF with Proximal Policy Optimization PPO or, more simply, via Direct Preference Optimization DPO.
arXiv:2605.00245v1 Announce Type: new Abstract: Large language models LLMs are now being explored for defense applications that require reliable and legally compliant decision support.
arXiv:2605.00248v1 Announce Type: new Abstract: A key challenge for the safety of advanced AI systems is the possibility that multiple simpler agents might inadvertently form a collective agent with capabilities and goals distinct from those of any individual.
arXiv:2605.00276v1 Announce Type: new Abstract: Trip planning for intelligent vehicles increasingly requires selecting optimal routes rather than merely producing feasible itineraries, as interacting factors such as travel time, energy consumption, and traffic conditions directly affect plan...
arXiv:2605.00300v1 Announce Type: new Abstract: Public inference benchmarks compare AI systems at the model and provider level, but the unit at which deployment decisions are actually made is the endpoint: the provider, model, stockkeepingunit tuple at which a specific quantization, decoding...
arXiv:2605.00334v1 Announce Type: new Abstract: Production agentic systems make many model calls per user request, and most of those calls are short, structured, and routine.
How OpenAI rebuilt its WebRTC stack to power realtime Voice AI with low latency, global scale, and seamless conversational turntaking.
사모펀드Private Equity, PE가 인공지능AI을 중심으로 근본적인 전환 국면에 진입하고 있다. 단순한 기술 도입을 넘어 기업 운영 방식과 가치 창출 구조 자체를 재설계하는 흐름이 본격화되면서, 엔터프라이즈 AI 경쟁이 산업 전반을 재편하고 있다.IBM 컨설팅 아메리카스 수석 부사장인 닐 다르Neil Dhar는 지난 1일현지시간 기고를 통해 "현재 경쟁에서 앞서 나가는 기업들은 특정 모델 하나에 의존하는 조직이 아니라 비즈니스 운영 방식을 재설계하고, 하이브리드 아키텍처를 구축하며, 시간이 지날수록 가치가 누적
The ad comes from Artisan, the AI startup behind billboards urging businesses to "stop hiring humans."
뉴럴링크의 뇌컴퓨터 인터페이스BCI 기술이 디지털 공간을 넘어 현실 세계로 확장됐다. 칩을 이식받은 환자가 생각만으로 로봇 팔을 움직이고 드론을 조종할 수 있게 됐다. 뉴럴링크 임플란트의 두번째 시술자인 알렉스 콘리는 최근 케이티 파블리치 투나잇과의 단독 인터뷰에서 뇌에 칩을 이식한 후 로봇 손으로 물건을 만들고 생각만으로 드론을 조종할 수 있게 된 과정에 대해 이야기했다.콘리는 2021년 차량 전복 사고로 척수 손상을 입어 하반신 마비가 됐고 휠체어에 의존하게 됐다. 그리고 2025년 7월 뉴럴링크의 칩을 두번째로 이식받았다