Everything the newsroom published, in chronological order. Each item carries origin, sources and reading time.
RESEARCH Schema.org Metadata Cuts Agentic Retrieval Errors by Two-Thirds
RESEARCH Stanford Framework Keeps AI Agents Within Violation Targets
RESEARCH Bidirectional Evolutionary Search Escapes Autoregressive Limits in Reasoning
RESEARCH MATCHA Outperforms BERTScore by 20% at Detecting Semantic Contradictions
RESEARCH BRANE Cuts Retrieval Agent Costs by 89% Per Query
RESEARCH RLHF Training Amplifies Model Bias to 100 Percent
RESEARCH Meta Shrinks Mixture-of-Experts to Smartphones Without Cloud Offloading
RESEARCH Mistral's 30B mixture-of-depths model remains unconfirmed but would fill a code-stack gap
RESEARCH ActiveGraph Inverts Agent Architecture, Putting Event Log First
RESEARCH LoopMDM Cuts Training FLOPs 3.3× by Recycling Transformer Layers
RESEARCH VeriTrace Improves Research Agents Without Scaling Models
RESEARCH Claw-Anything Benchmark Sets 34.5% Ceiling for Always-On Agents
RESEARCH OrpQuant Runs 7B Models on Edge Silicon Without Multipliers
RESEARCH IBM Framework Classifies Code Changes at 84% Recall
RESEARCH Self-Generated Replay Cuts Catastrophic Forgetting in Fine-Tuned Models
RESEARCH Stanford Framework Reveals Hidden Flaws in AI Benchmarks
RESEARCH MobileGym Solves Mobile-Agent Reproducibility at Scale
RESEARCH Study: AI Narrative Explanations Boost User Trust, Not Accuracy
RESEARCH Model Scale Fails to Predict Extracted Skill Performance
RESEARCH Five Bugs Killed agentmemory in Seven Days
RESEARCH Shannon-Hartley Theorem Explains LLM Quantization Regressions
RESEARCH Complete-muE Lets Teams Transfer Dense Hyperparameters to MoE
RESEARCH Microsoft's SkillOpt Lifts Agent Accuracy 24 Points via Automated Skill Refinement
RESEARCH MemAudit Cuts Memory-Poisoning Attacks to 0%
RESEARCH Six Chatbots Show 12-Point Accuracy Drop on Hindi News
RESEARCH Matching Principle Unifies Seven Robustness Families
RESEARCH Gated DeltaNet-2 Beats Linear Baselines on Long-Context Retrieval
RESEARCH Vector Policy Optimization beats GRPO on diverse sampling
RESEARCH Self-Modifying Agents Boost Benchmark Score to 0.61
RESEARCH DeltaBox cuts AI agent checkpoint latency to 14 milliseconds
RESEARCH LCGuard Patches KV-Cache Leakage in Multi-Agent Systems
RESEARCH DelTA Framework Improves Reasoning by Fixing Token-Level Credit Assignment
RESEARCH Equilibrium Reasoners lift Sudoku accuracy from 2.6% to 99% via test-time scaling
RESEARCH NVIDIA's CARV cuts 3D distillation compute by 2–3×
RESEARCH One hyperparameter rule captures most of µP's gains
RESEARCH RELEX reconstructs RLVR checkpoints from 15% training data
RESEARCH Peking researchers release DeepWeb-Bench, exposing derivation failures in frontier AI
RESEARCH Fine-tuning erases reasoning chains while accuracy stays high