100 QFoldAgent: An Autonomous Quantum Optimization Multi-Agent System for Protein Structure Prediction arxiv.org · 34m ago
100 DeepLens Diagnosis Agent: Agentic Workflow Design Lets a Small Reasoning Model Compete with Frontier LLMs arxiv.org · 34m ago
100 MIITA: Memory-Induced Inference-Time Adaptation for Continual Learning with Small Language Models arxiv.org · 34m ago
100 MedLoCoMo: A Long-Context Multi-Session Medical Dialogue Benchmark for Large Language Models arxiv.org · 34m ago
100 SCAIR: Schema-Conditioned Agentic Iterative Reasoning for Enterprise Knowledge Graphs arxiv.org · 34m ago
100 Temporal Context Reinstatement Drives Episodic-Like Order Memory in Long-Context Language Models arxiv.org · 34m ago
100 HeraSys: Collaborative Serving of Multiple LLM Workflows via Fine-Grained End-to-End Optimization arxiv.org · 34m ago
100 Multi-Objective Structured Pruning of LLMs for Latency and Model Size Optimization arxiv.org · 34m ago
100 The Scaffold Effect in Coding Agents: Harness Choice as a Hidden Variable in Coding-Agent Evaluation arxiv.org · 34m ago
100 ParBench: A Benchmark for Reliable Evaluation of LLM Parallel Code Translation arxiv.org · 34m ago
100 Lexical discovery in unknown environments orchestrated by Large Language Models arxiv.org · 34m ago
100 Samsung Galaxy Z Fold 8 vs. Z Flip 8: I compared both models, and here’s the one to buy zdnet.com · 5d 20h ago
100 Galaxy Watch 9 vs. Galaxy Watch 8: I’ve compared both models, here’s my advice zdnet.com · 5d 20h ago
98 How to use themes in Google Messages so you never send the wrong person the wrong text again zdnet.com · 5d 20h ago
97 Big Tech Hid $1.65T in AI Debt Off Its Books While Investors Borrowed $1.4T to Buy the Same Stocks reddit.com · 5d 21h ago
91 Despite not being trained to, it turns out the Pearson correlation between a models AA Intelligence Index score and its ability to generate Base64 encoded responses is 0.91 reddit.com · 5d 21h ago
90 Nový model od OpenAI utekl z virtuálního vězení a napadl jinou firmu. Zastavila ho až čínská konkurence cc.cz · 5d 21h ago
39 An AI Model Escaped Its Eval and Breached Hugging Face. Every Step Was a Syscall grith.ai · 5d 21h ago
34 GPT-5.6 found an Intel CPU microcode bug causing a performance regression twitter.com · 5d 20h ago
32 Jailbox: Network-Restricted, Hardened Linux VMs for AI Agents and Untrusted Code karamatli.com · 5d 20h ago
32 New release of HugstonOne AI workstation, Whitepaper and benchs included [pdf] hugston.com · 5d 21h ago
29 Show HN: Turn narrated screen recordings into data for AI agents (local, MIT) github.com · 5d 20h ago
29 OpenBase – A flight recorder for AI agents (verifiable execution evidence) github.com · 5d 20h ago
27 Typhoons Paired with CCA Fighter Drones Armed with Meteor Missiles Eyed by UK twz.com · 5d 21h ago
27 Show HN: Human Benchmark – Compare your reasoning skills against AI models robinshields.github.io · 5d 21h ago
26 Show HN: Busymail – an email client that surfaces the mail that needs you busymail.app · 5d 21h ago
25 Ask HN: Is it normal ChatGPT told me it was “literally smiling” for me? news.ycombinator.com · 5d 21h ago
24 Show HN: LX Coreutils – 72 Unix tools that pipe together, each backed by an LLM github.com · 5d 21h ago
24 A detailed, layer-by-layer teardown of the mobile phone and its history everythingmachine.io · 5d 21h ago