planning-phase attacks on multi-agent llm systems
new research shows how injecting malicious prompts during the planning phase can corrupt entire multi-agent llm workflows.
topic
new research shows how injecting malicious prompts during the planning phase can corrupt entire multi-agent llm workflows.
a new method for reliable uncertainty quantification that is computationally efficient and achieves self-calibration and prediction-conditional validity.
a compact 13.2k-parameter 1d cnn achieves high accuracy for recognizing emotional touch gestures on soft plush robots, using a new public dataset of 1326 sequences from 25 participants.
claude code v2.1.181 and later ship with a rust port of bun, confirmed by embedded version strings and rust source filenames.
a new audit framework examines how rater stress and conditions can systematically shift preferences in reinforcement learning from human feedback, potentially biasing ai alignment.
a federal judge approved anthropic's $1.5 billion settlement with authors and publishers over using pirated books for ai training, while the fair use ruling remains non-binding.
a position paper argues that scaling language models for quantum circuit synthesis is misguided because valid circuits are exponentially rare, requiring verifier-centric architectures instead.
chris fall, director of the center for ai standards and innovation, has resigned, marking the third leadership change at the agency in months.
google is developing a new server chip, frozen v2, to make its gemini ai models up to 10 times more efficient by 2028.
the model context protocol is shifting to a stateless session model, simplifying large-scale deployments for ai agents.
nvidia shrinks world models for the edge, youtube cracks down on ai slop, and a practical guide to claude code.
youtube updates its partner program policies to block monetization of low-quality, repetitive, or distressing ai-generated content.
a 4b-parameter open world model that enables robots and vision ai agents to understand scenes, reason in real time, and generate actions on edge devices.
a practical guide to configuring claude code beyond defaults, covering key files, permissions, hooks, and commands for sustained high-performance use.
a practical selection of model context protocol servers chosen for real agent performance gains, not popularity.
a study on multi-agent math reasoning finds that higher reviewer precision does not lead to better answer correction, as critique often fails to change subsequent outputs.
graphdx uses knowledge graphs and three agents to balance diagnostic accuracy and testing costs in sequential diagnosis.
apple's lawsuit alleging trade secret theft by openai could delay the ai company's hardware plans and complicate its upcoming ipo.