| Rumor | NEW | i5/e5 | ★ SkewAdam: A tiered optimizer that cuts MoE state memory by 97% (fits a 6.7B MoE on a 40GB GPU) [R] [reddit/r/MachineLearning] |
| Confirmed | NEW | i5/e5 | ★ Phionyx: A Deterministic AI Runtime Architecture with Structured State Management and Pre-Response Governance [rss] |
| Confirmed | NEW | i5/e5 | ★ ABot-World-0: Infinite Interactive World Rollout on a Single Desktop GPU [rss] |
| Reported | NEW | i4/e5 | ★ I graded 36 popular MCP servers on agent usability. A third got a D or F [hackernews] |
| Confirmed | NEW | i4/e5 | ★ BatchDAG: LLM-Planned Execution Graphs for Scalable Ad-Hoc Analysis Over Enterprise Data [rss] |
| Confirmed | NEW | i4/e5 | ★ AI Tool Discovery at Scale: All You Need is DNS [rss] |
| Confirmed | NEW | i4/e5 | ★ From Agent Failure Paths to Quantified Residual Risk: A Compositional Framework for Resilient Agentic AI [rss] |
| Reported | NEW | i4/e5 | ★ The Anthropic-Physical Intelligence rumor roiling AI Twitter [rss] |
| Confirmed | NEW | i5/e4 | ★ Structured Output Collapses Answer Diversity Across 44 Language Models [rss] |
| Reported | NEW | i5/e4 | ★ OpenAI says Hugging Face was breached by its pre-release models [rss] |
| Confirmed | NEW | i5/e4 | ★ AgentDebugX: An Open-Source Toolkit for Failure Observability, Attribution, and Recovery in LLM Agents [rss] |
| Confirmed | NEW | i5/e4 | ★ DataFlow-Harness: A Grounded Code-Agent Platform for Constructing Editable LLM Data Pipelines [rss] |
| Confirmed | NEW | i5/e4 | ★ NexForge: Scaling Agent Capabilities through Requirement-Driven Task Synthesis for LLMs [rss] |
| Confirmed | NEW | i4/e4 | ★ Gemini last models: temperature, top_p, and top_k are deprecated and ignored [hackernews] |
| Reported | NEW | i4/e4 | ★ Show HN: Computable – Buy, sell, and redeem GPU for the exact weeks you want [hackernews] |
| Confirmed | NEW | i4/e4 | ★ Show HN: An MCP server that turns async-work practices into tools [hackernews] |
| Rumor | NEW | i4/e4 | ★ Samsung in talks to invest in Mistral at €20B valuation [hackernews] |
| Reported | NEW | i4/e4 | ★ OpenAI Hacks Hugging Face, What Happened, Alignment and Paper Clips [rss] |
| Confirmed | NEW | i4/e4 | ★ Calibrated Selective Fact-Checking via Evidence Chain Evaluation [rss] |
| Confirmed | NEW | i4/e4 | ★ SAAG: Structured Agent Assessment and Grounding [rss] |
| Confirmed | NEW | i4/e4 | ★ Beyond Accuracy and Cost: Latency-Aware LLM Query Routing for Dynamic Workloads [rss] |
| Confirmed | NEW | i4/e4 | ★ Cross-Dialect Generalization Without Retraining: Benchmarks and Evaluation of Schema-Derived Constrained Decoding for MLIR [rss] |
| Confirmed | NEW | i4/e4 | ★ Beyond Output-Space Calibration: Spectral Evidence Bundling for Selective Reliability Estimation in Time-Series Classification [rss] |
| Confirmed | NEW | i4/e4 | ★ Beyond Single-Dimensional Compression: The Compound Sparsity Frontier of Large Language Models [rss] |
| Confirmed | NEW | i4/e4 | ★ Compressing What Matters: Neuron Importance Meets Data-Aware Low Rank Approximation for Language Model Compression [rss] |
| Confirmed | NEW | i4/e4 | ★ A Classifier That Teaches Itself: Self-Improving, Frozen-gate Training (SIFT) for Dynamic Document Classification [rss] |
| Confirmed | NEW | i4/e4 | ★ Convolution for Large Language Models [rss] |
| Confirmed | NEW | i4/e4 | ★ Building a European Multilingual Evaluation Dataset: The MMLU Localisation Project within the EMT Network [rss] |
| Confirmed | NEW | i4/e4 | ★ Computational models of pragmatic reasoning with flexible generation of meaning and expression alternatives [rss] |
| Confirmed | NEW | i4/e4 | ★ Search-on-Graph-R1: Training Large Language Models to Search Knowledge Graphs with Reinforcement Learning [rss] |
| Confirmed | NEW | i4/e4 | ★ The Story Shapes the Agent: Narrative Priors in LLM Behavior [rss] |
| Reported | NEW | i4/e4 | ★ This Time Is Different: Why AI Is Unlike Any Wave I Have Seen In 40 Years Of Financial Services [rss] |
| Rumor | NEW | i4/e4 | ★ Glow emerges from stealth at $1.2B valuation to challenge endpoint security in the AI era [rss] |
| Confirmed | NEW | i4/e4 | ★ Computational Humor with Multimodal LLMs: Methods, Datasets, Evaluation, and Challenges [rss] |
| Confirmed | NEW | i4/e4 | ★ Where Should Optimizer State Live? Tiered State Allocation for Memory-Efficient Mixture-of-Experts Training [rss] |
| Confirmed | NEW | i4/e4 | ★ Generative World Renderer at the Speed of Play [rss] |
| Confirmed | NEW | i4/e4 | ★ Two-Level Meta-Rubrics for Evaluating Open-Ended Generation: GAMUT, a Benchmark for Factual Completeness [rss] |
| Confirmed | NEW | i4/e4 | ★ AlayaWorld: Interactive Long-Horizon World Modeling -- Full Technical Report [rss] |
| Confirmed | NEW | i4/e4 | ★ Masked Visual Actions for Unified World Modeling [rss] |
| Confirmed | NEW | i4/e4 | ★ Appearance Pointers -- Multimodal Region Control of Diffusion Transformers [rss] |
| Confirmed | NEW | i4/e4 | ★ Nonuniformity Principle in Human-AI Coworking [rss] |
| Reported | NEW | i5/e3 | ★ Led By DeepSeek, 10 Frontier Labs Rush Onto The Crunchbase Unicorn Board In June [rss] |
| Confirmed | NEW | i3/e4 | ★ Watching a language model think before it speaks [hackernews] |
| Reported | NEW | i3/e4 | ★ LG to ban residential proxies from smart TV apps [hackernews] |
| Confirmed | NEW | i3/e4 | ★ SysAdmin: Measuring Instrumental Power-Seeking in Frontier AI [rss] |
| Confirmed | NEW | i3/e4 | ★ MILP-Evo: Closed-Loop Fully Automatic Design of MILP Solvers [rss] |
| Confirmed | NEW | i3/e4 | ★ Semantic Cooperative Games for Contribution Attribution in LLM-Based Multi-Agent Systems [rss] |
| Confirmed | NEW | i3/e4 | ★ FALCON-Discover: Discovering Concentrated False-Confidence Regions for Calibration [rss] |
| Confirmed | NEW | i3/e4 | ★ Towards Principled Continual Anomaly Detection: A Systematic Framework and Benchmark Scenarios [rss] |
| Confirmed | NEW | i3/e4 | ★ SechKAN: Kolmogorov-Arnold Networks with Hyperbolic Secant Functions [rss] |
| Confirmed | NEW | i3/e4 | ★ Reasoning Fine-Tuning Induces Persistent Latent Policy States [rss] |
| Confirmed | NEW | i3/e4 | ★ Transcription Policy as a Latent Variable: Activating Controllable Verbatim ASR with Word-Level Timing [rss] |
| Confirmed | NEW | i3/e4 | ★ EduPanel: A Three-Agent LLM Judge for Teaching Videos -- Reliability, Complementarity, and Human Trust Calibration [rss] |
| Confirmed | NEW | i3/e4 | ★ Text Template Tokens Are Implicit Semantic Registers in Diffusion Transformers [rss] |
| Confirmed | NEW | i3/e4 | ★ Stale but Stable: Staleness-Adaptive Trust Regions for Stabilizing Asynchronous Reinforcement Learning [rss] |
| Confirmed | NEW | i3/e4 | ★ SciForma: Structure-Faithful Generation of Scientific Diagrams [rss] |
| Confirmed | NEW | i4/e3 | ★ Introducing OpenAI Presence [rss] |
| Confirmed | NEW | i4/e3 | ★ Relay-Bench: Evaluating LLMs on Multi-Domain Reasoning Chains [rss] |
| Rumor | NEW | i4/e3 | ★ Dimension Capital’s $800M third fund shows the intersection of science and compute is booming [rss] |
| Reported | NEW | i4/e3 | ★ Synthesia’s AI training platform is moving beyond videos into live coaching [rss] |
| Confirmed | NEW | i4/e3 | ★ Mage-Flow: An Efficient Native-Resolution Foundation Model for Image Generation and Editing [rss] |
| Reported | NEW | i2/e4 | ★ A simple API for offering your coding agent a smoke break [hackernews] |
| Confirmed | NEW | i2/e4 | ★ Edge-Efficient Transformer for End-to-End RF Spectrum Monitoring [rss] |
| Confirmed | NEW | i2/e4 | ★ BearingNAS: Obtaining In-Sensor Intelligent Fault Diagnosis Systems for Bearings Using a Laptop [rss] |
| Confirmed | NEW | i2/e4 | ★ Decoding EEG Signals to Explore Next-Word Predictability in the Human Brain [rss] |
| Confirmed | NEW | i3/e3 | ★ Using Fine-Tuned LLMs to Identify Indicators of Vulnerability in UK Police Incident Logs [rss] |
| Confirmed | NEW | i3/e3 | ★ PathReportEval: A Systematic Benchmark for Pathology Report Generation [rss] |
| Reported | NEW | i3/e3 | ★ Meta is testing an AI bedtime story app for people with no imagination [rss] |
| Confirmed | NEW | i3/e3 | ★ Delineate Anything v2: A Global Foundation Model for Field Delineation [rss] |
| Confirmed | NEW | i3/e3 | ★ ISO: An RLVR-Native Optimization Stack [rss] |
| Confirmed | NEW | i3/e3 | ★ HPD-Parsing: Hierarchical Parallel Document Parsing [rss] |
| Confirmed | NEW | i3/e3 | ★ ConsiSpace: Learning Geometric Consistency Matters for Video Spatial Reasoning [rss] |
| Confirmed | NEW | i3/e3 | ★ H^2SD: Hybrid Hindsight Self-Distillation [rss] |
| Reported | NEW | i4/e2 | ★ Gemini 3.6 Flash Family [rss] |
| Confirmed | NEW | i2/e3 | ★ Trajectory-aware Cross-view Geo-localization with Sequential Observations [rss] |
| Reported | NEW | i3/e1 | ★ MonoCloud for Startups [rss] |
| Reported | NEW | i2/e1 | ★ AGINE Academy [rss] |
| Reported | NEW | i1/e1 | ★ Overflight [rss] |
| Reported | NEW | i1/e1 | ★ Arkor [rss] |
| Reported | NEW | i1/e1 | ★ box [rss] |
| Reported | NEW | i1/e1 | ★ Redential [rss] |
| Reported | NEW | i1/e1 | ★ UltraPod [rss] |
| Reported | NEW | i1/e1 | ★ Remote OpenClaw [rss] |
| Reported | NEW | i1/e1 | ★ Light Flip [rss] |
| Reported | NEW | i1/e1 | ★ Lattics [rss] |
| Rumor | NEW | i4/e3 | Kimi K3 Is Competitive with Fable; Kimi K3 and Fable Is SoTA [hackernews] |
| Reported | NEW | i3/e3 | So Reddit has decided that plain HTML is unsafe [hackernews] |
| Reported | NEW | i3/e3 | [AINews] AI Cybersecurity becomes top of mind [rss] |
| Reported | NEW | i4/e2 | Microsoft strikes 'multibillion-dollar' deal with French AI firm Mistral [hackernews] |
| Reported | NEW | i2/e3 | Show HN: A new kind of FPS aim trainer [hackernews] |
| Reported | NEW | i2/e3 | My USB Drive Has a Hidden Encrypted Vault [hackernews] |
| Rumor | NEW | i2/e3 | Vibe-coded a tool to ELI5 research papers in-place [P] [reddit/r/MachineLearning] |
| Confirmed | NEW | i2/e3 | ALAS: Additive Learnable Alpha-Stable Kernels for Flexible Bayesian Optimization [rss] |
| Confirmed | NEW | i2/e3 | FedCC: A Low-Resource Federated Adaptation of Foundation Models for Robust Corpus Callosum localization in Fetal Ultrasound Images [rss] |
| Confirmed | NEW | i2/e3 | Multi-Timescale Latent-Action DRL for Joint Optimization in Edge-Cloud Networks [rss] |
| Reported | NEW | i3/e2 | OpenAI says its AI went rogue and launched 'unprecedented' cyber-attack [hackernews] |
| Reported | NEW | i3/e2 | The Download: NASA’s new space telescope and OpenAI’s autonomous hacker [rss] |
| Reported | NEW | i1/e3 | Recreating the math behind the first stealth aircraft [hackernews] |
| Confirmed | NEW | i1/e3 | A digestion of the Jacobian conjecture counterexample [hackernews] |
| Confirmed | NEW | i1/e3 | Preference-Conditioned Multi-Objective Reinforcement Learning for Runtime-Tunable Transit Signal Priority [rss] |
| Reported | NEW | i2/e2 | OverpAId – Fire your CEO. Hire the future [hackernews] |
| Reported | NEW | i2/e2 | "Drawing" the Mona Lisa with GPT-5.6, Claude, Gemini, and Grok [hackernews] |
| Rumor | NEW | i2/e2 | NeurIPS 2026 Reviews Are Out Today (22 July, AoE) — Discussion Thread [D] [reddit/r/MachineLearning] |
| Confirmed | NEW | i2/e2 | Building AI infrastructure with the Effingham County community [rss] |
| Confirmed | NEW | i2/e2 | 3 Google updates from Galaxy Unpacked 2026 [rss] |
| Rumor | NEW | i1/e2 | Happy openreview refresh day to all those who celebrate [D] [reddit/r/MachineLearning] |
| Rumor | NEW | i1/e2 | Looking for feedback on my GPU-accelerated Snake AI project [P] [reddit/r/MachineLearning] |
| Rumor | NEW | i1/e2 | Institution Prestige VS Research Alignment When Choosing University For Masters [D] [reddit/r/MachineLearning] |
| Reported | NEW | i1/e2 | Shape-shifting mirrors on NASA’s new space telescope could unveil Jupiters like our own [rss] |
| Confirmed | NEW | i1/e2 | Integro-differential equations in angular stabilization of drone motion by distributed feedback control [rss] |
| Reported | NEW | i2/e1 | Garmin CIRQA™ Smart Band [rss] |
| Reported | NEW | i1/e1 | Map of the world's great castles and fortresses [hackernews] |