100 Turn your ordinary bike into an indoor trainer with this Wi-Fi connected smart base theverge.com · 1h 39m ago
98 GM redesigned its engineering workflows around AI agents — and tripled its merged pull requests venturebeat.com · 58m ago
97 Runway couldn’t fix a bug in its AI video model, so it turned the bug into a feature venturebeat.com · 58m ago
97 Invibes advertsing : Invibes Advertising recule au premier semestre mais accélère le déploiement de Fusion dans la Connected TV tradingsat.com · 1h 23m ago
100 QFoldAgent: An Autonomous Quantum Optimization Multi-Agent System for Protein Structure Prediction arxiv.org · 10h 16m ago
100 Same Question, Different Answers: Evaluating LLM Reliability Beyond Accuracy arxiv.org · 10h 16m ago
100 DeepLens Diagnosis Agent: Agentic Workflow Design Lets a Small Reasoning Model Compete with Frontier LLMs arxiv.org · 10h 16m ago
100 MIITA: Memory-Induced Inference-Time Adaptation for Continual Learning with Small Language Models arxiv.org · 10h 16m ago
100 MedLoCoMo: A Long-Context Multi-Session Medical Dialogue Benchmark for Large Language Models arxiv.org · 10h 16m ago
100 Keyword Matters: Unveiling the Energy Sensitivity of On-Device LLM Prompting arxiv.org · 10h 16m ago
100 SCAIR: Schema-Conditioned Agentic Iterative Reasoning for Enterprise Knowledge Graphs arxiv.org · 10h 16m ago
100 Temporal Context Reinstatement Drives Episodic-Like Order Memory in Long-Context Language Models arxiv.org · 10h 16m ago
100 VeriSimpl: Robust Optimization Modeling from Natural Language using Simplification-based Verification arxiv.org · 4d 13h ago
100 Benchmarking Large Language Models on Multi-Sensor Physical Hazard Assessment arxiv.org · 4d 13h ago
100 Beyond Liars’ Bench: The Impact of Lie Typology, Depth, and Sparsity on Deception Detection in LLMs arxiv.org · 4d 13h ago
100 Enabling Scalable Topology Inference in Distribution Systems via Constrained Multi-Source Inference arxiv.org · 4d 13h ago
100 Routing Without Training: Controllable-Ratio LLM Offloading via Reliability Gating arxiv.org · 4d 13h ago
100 PersonaTrail: Benchmarking Personalized Web Agents through Browsing Trails arxiv.org · 4d 13h ago
100 The Devil is in the Spectrum: Mitigating Representation Collapse in LLMs via Topologically Regularized Side-Path arxiv.org · 4d 13h ago
100 OPTScientist: Multi-Agent Discovery of Typed Optimizer Programs for Transformer Pretraining arxiv.org · 4d 13h ago
100 Directional Hallucinations: Ideological Drift in News-Grounded LLM Question Answering arxiv.org · 4d 13h ago
100 Autonomous Topology Mutation: Safe Runtime Restructuring for Multi-Agent LLM Systems with Capability, State, and Shadow Invariants arxiv.org · 4d 13h ago
100 DFAH-Bench: Benchmarking Observable Agent Instability in Financial Decision-Making arxiv.org · 4d 13h ago
69 If MOEs have small experts (3B/4B/9B etc), then why can’t we have small expert models as a whole rather than one large model with multiple experts? Like Qwen3.6-3B Coding Expert or something reddit.com · 4d 15h ago
55 Codex with GPT 5.6 Sol Ultra is a powerhouse, and doing things i never thought possible this early. reddit.com · 4d 15h ago
47 23-Year-Old AI Billionaire Says Gen Z Should Avoid This Popular Career Tactic entrepreneur.com · 4d 16h ago
43 The Marketing Skill Nobody Trains You On (And It’s Quietly Killing Your Deadlines) entrepreneur.com · 4d 16h ago
34 Businesses Are Making a Costly AI Mistake , and It Isnt What You Think prnewswire.com · 4d 15h ago
29 Show HN: Deep Skill Finder – Find agent skills using real execution benchmarks meyo.life · 4d 16h ago
28 AI Was Supposed to Lift Everybody., The Price Tag Says Otherwise allyagentoperations.com · 4d 16h ago
26 In the 1950s, a Soviet engineer proposed damming the Bering Strait economictimes.indiatimes.com · 4d 15h ago
25 AI will not trigger employment collapse, staffing company Adecco Group says reuters.com · 4d 15h ago
23 Seeing Is Not Believing: Realistic AI Videos Disrupt Confidence in Real Videos media.mit.edu · 4d 15h ago
19 Execution-Free Agentic Program Repair for Enterprise-Scale Development [pdf] dl.acm.org · 4d 13h ago
19 Uncle Bob: My current strategy is to not read any code written by my agents xcancel.com · 4d 14h ago