agentwall adds runtime safety for local ai agents
agentwall intercepts agent actions before execution, enforcing declarative safety policies to prevent harmful operations on local machines.
aisummaries filed under research
agentwall intercepts agent actions before execution, enforcing declarative safety policies to prevent harmful operations on local machines.
aiagentstop reduces energy use in local ai agents by predicting task failure early, cutting token waste and battery drain on consumer devices.
aia new method learns safe decision rules from logged data using general risk measures like cvar, with strong theoretical guarantees.
aia lightning talk summary of major llm developments from november 2025 to may 2026, including coding agent breakthroughs and open-weight model advances.
aisandboxaq integrates its physics-grounded quantitative models into anthropic's claude, letting researchers run complex simulations through plain language prompts.
aiinstruction-tuned language models show fair outputs but retain biased internal representations that can reverse decisions when activated.
aia trust-region method for fine-tuning multi-agent llm teams avoids compounding errors from stale rollouts, outperforming baselines by 7.1%.
aia new open benchmark evaluates complete agent systems, not just models, across six diverse tasks to measure generality, quality, and cost.
aideepslide is a multi-agent system that helps prepare entire presentations, from planning and slide creation to rehearsal support.
aia new method learns the optimal price as a smooth function of a scalar index, improving revenue in semiparametric dynamic pricing.
aiimproving llm theory of mind on static tests often fails to boost real interactive task performance, a new study finds.
aia new framework uses state machines and online rlhf to enforce business rules in multi-agent systems, improving accuracy over gpt-4o on a recruitment benchmark.
aia new framework replaces the unstable indicator function in tcav with a smooth parameterized function, reducing variance and unifying existing methods.
aia study finds that aggressive quantization causes previously unbiased language models to develop new stereotypical behaviors, with a clear dose-response pattern.
aithe uk government digital service recommends keeping public sector code open by default, pushing back against the nhs decision to close repositories after vulnerability reports.
aiarxiv will impose a one-year ban on authors who submit papers with clear signs of unverified llm output, such as hallucinated references.
aia new mixed integer goal programming method creates personalized meal plans using whole servings and soft nutrient targets, avoiding fractional foods and infeasible diets.
aigoogle's turboquant uses polarquant and qjl to cut kv cache memory by over 5x without retraining or accuracy loss.
ainasa is testing a radiation-hardened processor that delivers up to 500 times the performance of current spaceflight computers, enabling ai-powered spacecraft to operate independently in deep space.
aia new weighted regret metric reveals deterministic online multiple testing procedures suffer linear regret from false negatives, and a decoupled wrapper fixes it.
ai