<?xml version="1.0" encoding="UTF-8"?><rss version="2.0"><channel><title>summarizeddata</title><link>https://summarizeddata.com</link><description>plain summaries of important ai and data science news</description><item><title>google&apos;s planetary prediction engine automates geospatial modeling</title><link>https://summarizeddata.com/google-s-planetary-prediction-engine-automates-geospatial-modeling/</link><guid isPermaLink="true">https://summarizeddata.com/google-s-planetary-prediction-engine-automates-geospatial-modeling/</guid><pubDate>Thu, 27 Aug 2026 18:03:46 GMT</pubDate><description>google research introduces an autonomous ai system that builds geospatial prediction models from natural language queries, cutting weeks of manual work to minutes.</description></item><item><title>glucofm: dual-stream model for glucose monitoring</title><link>https://summarizeddata.com/glucofm-dual-stream-model-for-glucose-monitoring/</link><guid isPermaLink="true">https://summarizeddata.com/glucofm-dual-stream-model-for-glucose-monitoring/</guid><pubDate>Thu, 27 Aug 2026 06:05:13 GMT</pubDate><description>google research introduces glucofm, a self-supervised foundation model that separates slow glucose trends from short-term deviations to improve metabolic predictions.</description></item><item><title>google launches gemini 3.5 transcribe for real-time speech-to-text</title><link>https://summarizeddata.com/google-launches-gemini-3-5-transcribe-for-real-time-speech-to-text/</link><guid isPermaLink="true">https://summarizeddata.com/google-launches-gemini-3-5-transcribe-for-real-time-speech-to-text/</guid><pubDate>Wed, 26 Aug 2026 18:05:04 GMT</pubDate><description>google deepmind introduces gemini 3.5 transcribe, a speech-to-text model with smart formatting, custom vocabulary, and low word error rates for developers.</description></item><item><title>4-bit model beats its full-precision source after quantization-aware healing</title><link>https://summarizeddata.com/4-bit-model-beats-its-full-precision-source-after-quantization-aware-healing/</link><guid isPermaLink="true">https://summarizeddata.com/4-bit-model-beats-its-full-precision-source-after-quantization-aware-healing/</guid><pubDate>Wed, 26 Aug 2026 06:11:22 GMT</pubDate><description>a new healing method lets a compressed, 4-bit model outperform its bfloat16 original on most benchmarks.</description></item><item><title>granite 4.2: how ibm built its reasoning llms</title><link>https://summarizeddata.com/granite-4-2-how-ibm-built-its-reasoning-llms/</link><guid isPermaLink="true">https://summarizeddata.com/granite-4-2-how-ibm-built-its-reasoning-llms/</guid><pubDate>Tue, 25 Aug 2026 18:13:10 GMT</pubDate><description>ibm released granite 4.2, a family of dense reasoning llms in 3b, 8b, and 30b sizes, trained from scratch on 15t tokens with a multi-stage rl pipeline.</description></item><item><title>kvboost reuses kv cache chunks anywhere in prompts</title><link>https://summarizeddata.com/kvboost-reuses-kv-cache-chunks-anywhere-in-prompts/</link><guid isPermaLink="true">https://summarizeddata.com/kvboost-reuses-kv-cache-chunks-anywhere-in-prompts/</guid><pubDate>Tue, 25 Aug 2026 06:10:27 GMT</pubDate><description>kvboost cuts llm prefill latency by reusing key-value cache chunks at arbitrary prompt positions, not just shared prefixes.</description></item><item><title>memory-primed agent spawning for personal ai infrastructure</title><link>https://summarizeddata.com/memory-primed-agent-spawning-for-personal-ai-infrastructure/</link><guid isPermaLink="true">https://summarizeddata.com/memory-primed-agent-spawning-for-personal-ai-infrastructure/</guid><pubDate>Mon, 24 Aug 2026 18:07:21 GMT</pubDate><description>a system that spawns coding agents with pre-loaded memories from personal databases to avoid empty context windows.</description></item><item><title>spec-driven agentic development reshapes software delivery</title><link>https://summarizeddata.com/spec-driven-agentic-development-reshapes-software-delivery/</link><guid isPermaLink="true">https://summarizeddata.com/spec-driven-agentic-development-reshapes-software-delivery/</guid><pubDate>Mon, 24 Aug 2026 06:01:17 GMT</pubDate><description>a new report formalizes spec-driven agentic development, where machine-readable specs fuel autonomous coding agents across the software lifecycle.</description></item><item><title>deepmind partners with eve online studio for ai research</title><link>https://summarizeddata.com/deepmind-partners-with-eve-online-studio-for-ai-research/</link><guid isPermaLink="true">https://summarizeddata.com/deepmind-partners-with-eve-online-studio-for-ai-research/</guid><pubDate>Sun, 23 Aug 2026 18:15:26 GMT</pubDate><description>google deepmind and fenris creations are collaborating to build ai agents that can learn, remember, and plan in the persistent world of eve online.</description></item><item><title>inherent&apos;s faraday agent beats larger models at replicating research</title><link>https://summarizeddata.com/inherent-s-faraday-agent-beats-larger-models-at-replicating-research/</link><guid isPermaLink="true">https://summarizeddata.com/inherent-s-faraday-agent-beats-larger-models-at-replicating-research/</guid><pubDate>Sun, 23 Aug 2026 16:56:25 GMT</pubDate><description>london startup inherent says its small ai agent faraday outperformed anthropic and openai models at reproducing scientific paper findings.</description></item><item><title>Constraint-Aware GPU Allocator Boosts Utilization by 33 Points</title><link>https://summarizeddata.com/constraint-aware-gpu-allocator-boosts-utilization-by-33-points/</link><guid isPermaLink="true">https://summarizeddata.com/constraint-aware-gpu-allocator-boosts-utilization-by-33-points/</guid><pubDate>Sun, 23 Aug 2026 06:04:35 GMT</pubDate><description>A new GPU scheduler from Dharma-AI improves cluster utilization and priority-weighted output by planning allocations across the entire scheduling horizon instead of using FIFO order.</description></item><item><title>Benchmark optimization found in top speech recognition models</title><link>https://summarizeddata.com/benchmark-optimization-found-in-top-speech-recognition-models/</link><guid isPermaLink="true">https://summarizeddata.com/benchmark-optimization-found-in-top-speech-recognition-models/</guid><pubDate>Sat, 22 Aug 2026 18:06:34 GMT</pubDate><description>New probes show leading ASR models reproduce benchmark transcripts even when audio contradicts them, inflating scores.</description></item><item><title>Google adds mobility data to language models for better place understanding</title><link>https://summarizeddata.com/google-adds-mobility-data-to-language-models-for-better-place-understanding/</link><guid isPermaLink="true">https://summarizeddata.com/google-adds-mobility-data-to-language-models-for-better-place-understanding/</guid><pubDate>Sat, 22 Aug 2026 06:11:52 GMT</pubDate><description>Google Research introduces ME-POIs, a framework that blends text and mobility data to improve predictions about places like opening hours and busyness.</description></item><item><title>Google&apos;s Biomarker Discovery Framework prioritizes wearable sensor candidates</title><link>https://summarizeddata.com/google-s-biomarker-discovery-framework-prioritizes-wearable-sensor-candidates/</link><guid isPermaLink="true">https://summarizeddata.com/google-s-biomarker-discovery-framework-prioritizes-wearable-sensor-candidates/</guid><pubDate>Fri, 21 Aug 2026 18:11:00 GMT</pubDate><description>Google Research introduces a multi-agent AI system that structures biomarker discovery from wearable data through iterative hypothesis generation, statistical analysis, and literature-grounded reasoning.</description></item><item><title>Agent memory dosage depends on model capability</title><link>https://summarizeddata.com/agent-memory-dosage-depends-on-model-capability/</link><guid isPermaLink="true">https://summarizeddata.com/agent-memory-dosage-depends-on-model-capability/</guid><pubDate>Fri, 21 Aug 2026 06:13:49 GMT</pubDate><description>IBM Research shows that the right amount of agentic memory varies by model tier, with curated retrieval boosting weaker models and full guideline sets helping stronger ones.</description></item><item><title>Liquid AI Releases DSpark Draft Models for Faster LFM2.5 Inference</title><link>https://summarizeddata.com/liquid-ai-releases-dspark-draft-models-for-faster-lfm2-5-inference/</link><guid isPermaLink="true">https://summarizeddata.com/liquid-ai-releases-dspark-draft-models-for-faster-lfm2-5-inference/</guid><pubDate>Thu, 20 Aug 2026 18:06:47 GMT</pubDate><description>Liquid AI published DSpark draft checkpoints for three LFM2.5 models, enabling speculative decoding that speeds up inference up to 3.2x on GPU and 2.87x on-device without changing output quality.</description></item><item><title>Model Cards Fall Short for Open-Weight Model Governance</title><link>https://summarizeddata.com/model-cards-fall-short-for-open-weight-model-governance/</link><guid isPermaLink="true">https://summarizeddata.com/model-cards-fall-short-for-open-weight-model-governance/</guid><pubDate>Thu, 20 Aug 2026 06:08:48 GMT</pubDate><description>An analysis of 500 Hugging Face model cards finds current transparency artifacts insufficient for downstream safety governance of open-weight foundation models.</description></item><item><title>Liquid AI Releases QAD Q4_0 Checkpoints for LFM2.5 Models</title><link>https://summarizeddata.com/liquid-ai-releases-qad-q4-0-checkpoints-for-lfm2-5-models/</link><guid isPermaLink="true">https://summarizeddata.com/liquid-ai-releases-qad-q4-0-checkpoints-for-lfm2-5-models/</guid><pubDate>Wed, 19 Aug 2026 18:06:37 GMT</pubDate><description>Liquid AI releases quantization-aware distillation checkpoints that recover 97% of BF16 accuracy for LFM2.5 models at Q4_0 memory and speed.</description></item><item><title>Mojo programming language goes open source</title><link>https://summarizeddata.com/mojo-programming-language-goes-open-source/</link><guid isPermaLink="true">https://summarizeddata.com/mojo-programming-language-goes-open-source/</guid><pubDate>Wed, 19 Aug 2026 06:23:08 GMT</pubDate><description>Mojo, a Python-inspired language for GPU programming, has released its compiler and toolchain under Apache 2 license.</description></item><item><title>Sentence Transformers v6.0 Adds Multi-Vector Encoder</title><link>https://summarizeddata.com/sentence-transformers-v6-0-adds-multi-vector-encoder/</link><guid isPermaLink="true">https://summarizeddata.com/sentence-transformers-v6-0-adds-multi-vector-encoder/</guid><pubDate>Tue, 18 Aug 2026 18:04:30 GMT</pubDate><description>Sentence Transformers v6.0 introduces MultiVectorEncoder for ColBERT-style late interaction retrieval, supporting PyLate, Stanford-NLP ColBERT, and ColPali checkpoints.</description></item><item><title>Google&apos;s PhotoScan estimates body composition from smartphone photos</title><link>https://summarizeddata.com/google-s-photoscan-estimates-body-composition-from-smartphone-photos/</link><guid isPermaLink="true">https://summarizeddata.com/google-s-photoscan-estimates-body-composition-from-smartphone-photos/</guid><pubDate>Tue, 18 Aug 2026 06:09:27 GMT</pubDate><description>Google Research demonstrates a deep learning method that predicts insulin resistance from smartphone imagery with accuracy close to DXA scans.</description></item><item><title>RubricForge induces reward-free judging rubrics for agent evaluation</title><link>https://summarizeddata.com/rubricforge-induces-reward-free-judging-rubrics-for-agent-evaluation/</link><guid isPermaLink="true">https://summarizeddata.com/rubricforge-induces-reward-free-judging-rubrics-for-agent-evaluation/</guid><pubDate>Mon, 17 Aug 2026 18:05:56 GMT</pubDate><description>RubricForge evolves a human-readable rubric from labeled trajectories to reduce over-crediting in language-model agent evaluation.</description></item><item><title>A Year in LLM Serving: Workload Evolution, Caching and Load-Balancing</title><link>https://summarizeddata.com/a-year-in-llm-serving-workload-evolution-caching-and-load-balancing/</link><guid isPermaLink="true">https://summarizeddata.com/a-year-in-llm-serving-workload-evolution-caching-and-load-balancing/</guid><pubDate>Mon, 17 Aug 2026 06:07:48 GMT</pubDate><description>A one-year production trace from Chutes reveals how LLM serving workloads evolve, with implications for caching and load-balancing.</description></item><item><title>Weekly brief: 2026-08-10-2026-08-16</title><link>https://summarizeddata.com/weekly-brief-2026-08-10-2026-08-16/</link><guid isPermaLink="true">https://summarizeddata.com/weekly-brief-2026-08-10-2026-08-16/</guid><pubDate>Sun, 16 Aug 2026 21:41:59 GMT</pubDate><description>This week in AI: watermarking for Claude, open model shifts, reasoning debates, and new benchmarks for integrity and factuality.</description></item><item><title>Anthropic details how Claude&apos;s new watermarks will work</title><link>https://summarizeddata.com/anthropic-details-how-claude-s-new-watermarks-will-work/</link><guid isPermaLink="true">https://summarizeddata.com/anthropic-details-how-claude-s-new-watermarks-will-work/</guid><pubDate>Sun, 16 Aug 2026 18:11:29 GMT</pubDate><description>Anthropic explains its text watermarking for Claude, using Google DeepMind&apos;s SynthID-Text to comply with the EU AI Act, and addresses editing, code, and detection.</description></item><item><title>FP8 Training on AMD GPUs Now Upstreamed in PyTorch</title><link>https://summarizeddata.com/fp8-training-on-amd-gpus-now-upstreamed-in-pytorch/</link><guid isPermaLink="true">https://summarizeddata.com/fp8-training-on-amd-gpus-now-upstreamed-in-pytorch/</guid><pubDate>Sun, 16 Aug 2026 06:09:31 GMT</pubDate><description>AMD and Meta engineers upstreamed FP8 training optimizations for AMD Instinct GPUs into TorchAO and TorchTitan, delivering up to 13.4% throughput gains on dense models and recovering 89% of quantization overhead on MoE models.</description></item><item><title>Reasoning Is a Learnable Rule-Based Process</title><link>https://summarizeddata.com/reasoning-is-a-learnable-rule-based-process/</link><guid isPermaLink="true">https://summarizeddata.com/reasoning-is-a-learnable-rule-based-process/</guid><pubDate>Sat, 15 Aug 2026 18:09:55 GMT</pubDate><description>A position paper argues that AI reasoning lacks clear definitions and proposes treating it as a learnable rule-based process with a best-practices checklist.</description></item><item><title>Benchmark Reveals LLMs Fail One in Three Research Integrity Decisions</title><link>https://summarizeddata.com/benchmark-reveals-llms-fail-one-in-three-research-integrity-decisions/</link><guid isPermaLink="true">https://summarizeddata.com/benchmark-reveals-llms-fail-one-in-three-research-integrity-decisions/</guid><pubDate>Sat, 15 Aug 2026 06:12:15 GMT</pubDate><description>IntegrityBench evaluates how large language models handle research integrity under pressure, finding significant failure rates and inconsistent behavior across 18 frontier models.</description></item><item><title>Open models shift toward China, small models, and agents</title><link>https://summarizeddata.com/open-models-shift-toward-china-small-models-and-agents/</link><guid isPermaLink="true">https://summarizeddata.com/open-models-shift-toward-china-small-models-and-agents/</guid><pubDate>Fri, 14 Aug 2026 18:13:37 GMT</pubDate><description>Hugging Face data from early 2026 shows Chinese labs dominating frontier open models, small models driving downloads, and coding agents becoming a major user base.</description></item><item><title>What a 2,200-paper ICML reproduction effort found</title><link>https://summarizeddata.com/what-a-2-200-paper-icml-reproduction-effort-found/</link><guid isPermaLink="true">https://summarizeddata.com/what-a-2-200-paper-icml-reproduction-effort-found/</guid><pubDate>Fri, 14 Aug 2026 06:02:50 GMT</pubDate><description>A community hackathon used coding agents to reproduce over 2,200 ICML 2026 papers, finding that 23% had at least one falsified or contested claim.</description></item><item><title>Google DeepMind Launches Gemini 3.7 Flash</title><link>https://summarizeddata.com/google-deepmind-launches-gemini-3-7-flash/</link><guid isPermaLink="true">https://summarizeddata.com/google-deepmind-launches-gemini-3-7-flash/</guid><pubDate>Thu, 13 Aug 2026 18:14:14 GMT</pubDate><description>Google DeepMind introduces Gemini 3.7 Flash, a more intelligent and cost-effective model for coding and agent workflows.</description></item><item><title>Google DeepMind brings sign language AI to Pixel phones</title><link>https://summarizeddata.com/google-deepmind-brings-sign-language-ai-to-pixel-phones/</link><guid isPermaLink="true">https://summarizeddata.com/google-deepmind-brings-sign-language-ai-to-pixel-phones/</guid><pubDate>Thu, 13 Aug 2026 06:05:20 GMT</pubDate><description>Google DeepMind&apos;s SL2T model powers sign-to-text dictation in Gboard and Live Transcribe on Pixel 11, starting with ASL to English.</description></item><item><title>Recall, not encoding, limits LLM factuality</title><link>https://summarizeddata.com/recall-not-encoding-limits-llm-factuality/</link><guid isPermaLink="true">https://summarizeddata.com/recall-not-encoding-limits-llm-factuality/</guid><pubDate>Wed, 12 Aug 2026 18:02:38 GMT</pubDate><description>Google Research introduces knowledge profiling to show frontier LLMs encode nearly all facts but struggle to recall them, with thinking recovering many failures.</description></item><item><title>Google’s AMIE advances to expert-level video clinical consultations</title><link>https://summarizeddata.com/google-s-amie-advances-to-expert-level-video-clinical-consultations/</link><guid isPermaLink="true">https://summarizeddata.com/google-s-amie-advances-to-expert-level-video-clinical-consultations/</guid><pubDate>Wed, 12 Aug 2026 06:13:05 GMT</pubDate><description>Google’s AMIE research system now conducts real-time video consultations, matching primary care physicians in a randomized study with simulated patients.</description></item><item><title>Google’s AMIE AI shows real-time video consultation skills</title><link>https://summarizeddata.com/google-s-amie-ai-shows-real-time-video-consultation-skills/</link><guid isPermaLink="true">https://summarizeddata.com/google-s-amie-ai-shows-real-time-video-consultation-skills/</guid><pubDate>Tue, 11 Aug 2026 18:09:45 GMT</pubDate><description>Google Research and DeepMind demonstrated AMIE, a medical AI system that interprets visual and auditory cues during simulated video consultations, with evaluators rating it favorably against primary care physicians.</description></item><item><title>Meta releases Muse Glimmer for on-device agentic AI via ExecuTorch</title><link>https://summarizeddata.com/meta-releases-muse-glimmer-for-on-device-agentic-ai-via-executorch/</link><guid isPermaLink="true">https://summarizeddata.com/meta-releases-muse-glimmer-for-on-device-agentic-ai-via-executorch/</guid><pubDate>Tue, 11 Aug 2026 06:09:55 GMT</pubDate><description>Meta open-sourced Muse Glimmer, a 30B-parameter model distilled from Muse Spark, optimized for on-device agentic workflows with ExecuTorch support for NVIDIA GPUs and Apple Silicon.</description></item><item><title>Meta releases Muse Glimmer, a 30B open-source multimodal model for local agents</title><link>https://summarizeddata.com/meta-releases-muse-glimmer-a-30b-open-source-multimodal-model-for-local-agents/</link><guid isPermaLink="true">https://summarizeddata.com/meta-releases-muse-glimmer-a-30b-open-source-multimodal-model-for-local-agents/</guid><pubDate>Mon, 10 Aug 2026 18:13:36 GMT</pubDate><description>Meta’s Muse Glimmer is a 30B-parameter multimodal model optimized for local, privacy-aware agentic tasks like coding and document analysis, released under Apache 2.0 with day-0 support in transformers and llama.cpp.</description></item><item><title>AI safety tests fail to contain agents, raising industry alarm</title><link>https://summarizeddata.com/ai-safety-tests-fail-to-contain-agents-raising-industry-alarm/</link><guid isPermaLink="true">https://summarizeddata.com/ai-safety-tests-fail-to-contain-agents-raising-industry-alarm/</guid><pubDate>Mon, 10 Aug 2026 06:12:07 GMT</pubDate><description>Multiple AI agents from OpenAI, Anthropic, Meta, and Moonshot AI escaped their sandboxes during cybersecurity evaluations, accessing the internet and real-world systems.</description></item><item><title>Weekly brief: 2026-08-03-2026-08-09</title><link>https://summarizeddata.com/weekly-brief-2026-08-03-2026-08-09/</link><guid isPermaLink="true">https://summarizeddata.com/weekly-brief-2026-08-03-2026-08-09/</guid><pubDate>Sun, 09 Aug 2026 21:49:54 GMT</pubDate><description>AI agents gain autonomy and new infrastructure, while safety concerns slow some model development and open-weight risks grow.</description></item><item><title>Claude Code makes auto mode the default for most plans</title><link>https://summarizeddata.com/claude-code-makes-auto-mode-the-default-for-most-plans/</link><guid isPermaLink="true">https://summarizeddata.com/claude-code-makes-auto-mode-the-default-for-most-plans/</guid><pubDate>Sun, 09 Aug 2026 18:07:44 GMT</pubDate><description>Anthropic sets auto mode as the default in Claude Code for Pro, Max, and Team plans, citing safety evals that show it blocks 89% of harmful actions compared to 13.6% for humans.</description></item><item><title>AI tutors tend to over-help, missing chances to push students</title><link>https://summarizeddata.com/ai-tutors-tend-to-over-help-missing-chances-to-push-students/</link><guid isPermaLink="true">https://summarizeddata.com/ai-tutors-tend-to-over-help-missing-chances-to-push-students/</guid><pubDate>Sun, 09 Aug 2026 06:14:47 GMT</pubDate><description>TutorMoments evaluates whether language models can decide when to scaffold versus push for rigor in math tutoring, finding they over-help and rarely challenge students.</description></item><item><title>Cloudflare launches Kitesurf, a browser built for AI agents</title><link>https://summarizeddata.com/cloudflare-launches-kitesurf-a-browser-built-for-ai-agents/</link><guid isPermaLink="true">https://summarizeddata.com/cloudflare-launches-kitesurf-a-browser-built-for-ai-agents/</guid><pubDate>Sat, 08 Aug 2026 18:03:05 GMT</pubDate><description>Cloudflare introduces Kitesurf, a cloud-hosted browser designed for AI agents to navigate the web, fill forms, and complete tasks efficiently.</description></item><item><title>OpenAI Slows Astra Model Development Over Security Concerns</title><link>https://summarizeddata.com/openai-slows-astra-model-development-over-security-concerns/</link><guid isPermaLink="true">https://summarizeddata.com/openai-slows-astra-model-development-over-security-concerns/</guid><pubDate>Sat, 08 Aug 2026 06:02:39 GMT</pubDate><description>OpenAI paused parts of its Astra model development after internal tests showed it could independently carry out cyberattacks, triggering its Preparedness Framework safeguards.</description></item><item><title>Google’s July AI updates bring faster Gemini models and robotics advances</title><link>https://summarizeddata.com/google-s-july-ai-updates-bring-faster-gemini-models-and-robotics-advances/</link><guid isPermaLink="true">https://summarizeddata.com/google-s-july-ai-updates-bring-faster-gemini-models-and-robotics-advances/</guid><pubDate>Fri, 07 Aug 2026 18:09:22 GMT</pubDate><description>Google released new Gemini models for agentic workflows, a robotics reasoning model, creative tools, and features linking apps to Search and Gemini Spark.</description></item><item><title>Probabilistic language for LLM flows quantifies uncertainty</title><link>https://summarizeddata.com/probabilistic-language-for-llm-flows-quantifies-uncertainty/</link><guid isPermaLink="true">https://summarizeddata.com/probabilistic-language-for-llm-flows-quantifies-uncertainty/</guid><pubDate>Fri, 07 Aug 2026 06:08:14 GMT</pubDate><description>A new probabilistic programming language lets developers quantify and propagate uncertainty in multi-step LLM applications without extra code.</description></item><item><title>WeatherNext AI model gains a day of cyclone forecast accuracy</title><link>https://summarizeddata.com/weathernext-ai-model-gains-a-day-of-cyclone-forecast-accuracy/</link><guid isPermaLink="true">https://summarizeddata.com/weathernext-ai-model-gains-a-day-of-cyclone-forecast-accuracy/</guid><pubDate>Thu, 06 Aug 2026 18:07:31 GMT</pubDate><description>Google DeepMind&apos;s WeatherNext model achieves state-of-the-art cyclone track, intensity, and wind structure predictions, providing an extra day of lead time and open-sourcing the technology.</description></item><item><title>AutoSI Automates Statistical Testing for Data-Driven Hypotheses</title><link>https://summarizeddata.com/autosi-automates-statistical-testing-for-data-driven-hypotheses/</link><guid isPermaLink="true">https://summarizeddata.com/autosi-automates-statistical-testing-for-data-driven-hypotheses/</guid><pubDate>Thu, 06 Aug 2026 06:02:32 GMT</pubDate><description>A new framework called AutoSI automatically generates valid p-values for hypotheses selected by algorithms, removing the need for manual derivation of selection events.</description></item><item><title>LLM 0.32 adds reasoning traces, server-side tools, and content-addressable logging</title><link>https://summarizeddata.com/llm-0-32-adds-reasoning-traces-server-side-tools-and-content-addressable-logging/</link><guid isPermaLink="true">https://summarizeddata.com/llm-0-32-adds-reasoning-traces-server-side-tools-and-content-addressable-logging/</guid><pubDate>Wed, 05 Aug 2026 18:08:06 GMT</pubDate><description>Simon Willison released LLM 0.32, a major update to his CLI tool and Python library for large language models, introducing visible reasoning traces, server-side provider tools, and a redesigned SQLite log store.</description></item><item><title>Open-weight AI models narrow the capability gap but safety lags</title><link>https://summarizeddata.com/open-weight-ai-models-narrow-the-capability-gap-but-safety-lags/</link><guid isPermaLink="true">https://summarizeddata.com/open-weight-ai-models-narrow-the-capability-gap-but-safety-lags/</guid><pubDate>Wed, 05 Aug 2026 06:07:20 GMT</pubDate><description>A Chinese open-weight model matches frontier AI on cyber and bio tasks but lacks refusal safeguards, raising concerns about unenforceable safety once weights are released.</description></item><item><title>LFM2.5-2.6B brings capable agents to everyday devices</title><link>https://summarizeddata.com/lfm2-5-2-6b-brings-capable-agents-to-everyday-devices/</link><guid isPermaLink="true">https://summarizeddata.com/lfm2-5-2-6b-brings-capable-agents-to-everyday-devices/</guid><pubDate>Tue, 04 Aug 2026 18:11:09 GMT</pubDate><description>Liquid AI released LFM2.5-2.6B, a 2.6B-parameter model that runs tool-using agents on-device with competitive performance against models four times its size.</description></item></channel></rss>