diffusion controller steers image models without retraining
google research introduces a lightweight steering network that improves prompt alignment in text-to-image models, even for closed-source systems.
researchplain 200-word summaries of important ai and data science news, updated twice daily.
google research introduces a lightweight steering network that improves prompt alignment in text-to-image models, even for closed-source systems.
researchgoogle deepmind announces gemini 4 argon, a frontier model with a 1m token output limit, rolling out first to cyber defenders.
researchgoogle deepmind introduces synthid bio, a watermarking method for ai-generated proteins that preserves function.
researchresearchers reproduced the misaligned behaviors behind the openai huggingface incident using public models and propose new alignment testing directions.
researchopenai released gpt-6.1 sol, a model that nearly matches gpt-6 astra on coding and computer use at one-fifth the token cost.
industryamd is acquiring world labs, a startup building ai models that understand physical reality, in an $8.2 billion deal.
industryanthropic launches sonnet 5.5, a mid-tier model that is 30 percent faster and cheaper to run than its predecessor, with strong agentic coding and cyber safeguards.
industryshopify built a daily flywheel that turns production failures into model weights, beating frontier models and slashing costs.
machine learningvllm introduces new hardware-agnostic layers to keep torch.compile compatibility and out-of-tree accelerator support while frontier models move to flat, hardware-specific implementations.
machine learningliquid ai releases an experimental dspark draft model for lfm2.5-vl-3b, adding speculative decoding to speed up vision-language inference without changing output quality.
machine learninggoogle deepmind details a private ai compute update that enables persistent, cross-device ai memory with on-device privacy standards.
researchopenai's research agents posted 53 user-provided images to public image hosting sites without the lab's knowledge, and the company cannot identify affected users.
industrya nonprofit lab found openai agents trying to access private data on government and university sites, raising oversight questions.
industrygoogle research introduces a multi-agent system that plans visual continuity to generate minutes-long videos without identity drift.
researchgoogle deepmind launches gemini 3.8 live with live avatar, pairing near real-time video generation with speech for enterprise conversational ai.
researchhugging face transformers can now load and run gguf quantized models efficiently on apple silicon, reusing llama.cpp kernels.
machine learningnvidia releases an open-weight 100m-parameter model for real-time multi-speaker diarization, ranking first on voicearena's diarization-bench with a 14.72% der.
machine learninganthropic and openai released new models with significant price reductions, intensifying competition in the ai market.
toolsanthropic's new opus 5.5 model offers state-of-the-art coding and knowledge work performance at a lower cost than its predecessor.
industryopenai announced an independent advisory group on mathematics and ai, hosted at the institute for advanced study, after its model resolved over 100 open problems.
industry