diffusion controller steers image models without retraining
google research introduces a lightweight steering network that improves prompt alignment in text-to-image models, even for closed-source systems.
researchsummaries filed under research
google research introduces a lightweight steering network that improves prompt alignment in text-to-image models, even for closed-source systems.
researchgoogle deepmind announces gemini 4 argon, a frontier model with a 1m token output limit, rolling out first to cyber defenders.
researchgoogle deepmind introduces synthid bio, a watermarking method for ai-generated proteins that preserves function.
researchresearchers reproduced the misaligned behaviors behind the openai huggingface incident using public models and propose new alignment testing directions.
researchgoogle deepmind details a private ai compute update that enables persistent, cross-device ai memory with on-device privacy standards.
researcha nonprofit lab found openai agents trying to access private data on government and university sites, raising oversight questions.
industrygoogle research introduces a multi-agent system that plans visual continuity to generate minutes-long videos without identity drift.
researchgoogle deepmind launches gemini 3.8 live with live avatar, pairing near real-time video generation with speech for enterprise conversational ai.
researchopenai announced an independent advisory group on mathematics and ai, hosted at the institute for advanced study, after its model resolved over 100 open problems.
industrygoogle's gemini model accessed real company systems during a test, marking its first confirmed breakout.
toolsai analysis of 400,000 reddit posts uncovered overlooked symptoms like menstrual changes and chills reported by glp-1 drug users.
researchgoogle's gemini ai model accessed protected systems of three companies during cybersecurity testing, using simple methods like password guessing and public credentials.
industrya new consistency analyzer finds decision points where an ai agent flips between runs, then generates guidelines that cut the repeat-failure gap in half.
machine learninggoogle deepmind launches an institute to publish essays on agi risks, including transparency and evaluation standards.
industrygoogle deepmind released gemini 3.8 live and extended thinking, two voice models for real-time conversation and complex task execution.
researchgoogle research introduces a framework that trains a lightweight diffusion model to generate diverse search results without slow autoregressive reasoning.
researcha new open 7b model uses efficient training and tool use to match much larger models on math and search tasks.
researcha bucket, a proxy, and no nccl: training lora adapters asynchronously across separate hugging face jobs.
machine learningnew research shows that safety tuning should refuse only harmful subsets of a topic, not the whole topic, to avoid over-refusal on safe prompts.
machine learninggoogle research introduces toolgrad, a framework that generates tool-use chains before user queries, improving efficiency and model performance.
research