100 QFoldAgent: An Autonomous Quantum Optimization Multi-Agent System for Protein Structure Prediction arxiv.org · 3h 5m ago
100 Same Question, Different Answers: Evaluating LLM Reliability Beyond Accuracy arxiv.org · 3h 5m ago
100 DeepLens Diagnosis Agent: Agentic Workflow Design Lets a Small Reasoning Model Compete with Frontier LLMs arxiv.org · 3h 5m ago
100 MIITA: Memory-Induced Inference-Time Adaptation for Continual Learning with Small Language Models arxiv.org · 3h 5m ago
100 MedLoCoMo: A Long-Context Multi-Session Medical Dialogue Benchmark for Large Language Models arxiv.org · 3h 5m ago
100 Keyword Matters: Unveiling the Energy Sensitivity of On-Device LLM Prompting arxiv.org · 3h 5m ago
100 SCAIR: Schema-Conditioned Agentic Iterative Reasoning for Enterprise Knowledge Graphs arxiv.org · 3h 5m ago
100 Temporal Context Reinstatement Drives Episodic-Like Order Memory in Long-Context Language Models arxiv.org · 3h 5m ago
100 HeraSys: Collaborative Serving of Multiple LLM Workflows via Fine-Grained End-to-End Optimization arxiv.org · 3h 5m ago
100 Multi-Objective Structured Pruning of LLMs for Latency and Model Size Optimization arxiv.org · 3h 5m ago
100 The Scaffold Effect in Coding Agents: Harness Choice as a Hidden Variable in Coding-Agent Evaluation arxiv.org · 3h 5m ago
100 ParBench: A Benchmark for Reliable Evaluation of LLM Parallel Code Translation arxiv.org · 3h 5m ago
97 As US weighs response to Chinese AI, industry urges against broad open-weight restrictions techcrunch.com · 3d 20h ago
97 The 3 types of people who will excel in the AI agent era, according to tech leaders zdnet.com · 3d 20h ago
95 I tested Dell’s new midrange work PC — It nails the sweet spot of price and performance zdnet.com · 3d 20h ago
91 Bluesky’s AI assistant Attie expands into an open social research tool techcrunch.com · 3d 20h ago
91 Need recommendations for small models with excellent reasoning. Professionals opinions preferred, this is for a data pipeline not chat. reddit.com · 3d 20h ago
82 First time seeing this notification while using 5.6 Sol extra high model in Chat reddit.com · 3d 20h ago
78 More than 20 companies including NVIDIA, Meta, Microsoft, Palantir, and Hugging Face have signed a letter urging policymakers to avoid premature restrictions on open weight models. reddit.com · 3d 20h ago
74 GPT-5.6 Thinking High surprised me on a 70+ page engineering compliance review — this felt very different from normal “PDF Q&A” reddit.com · 3d 20h ago
69 Jensen Huan created his official X account just to share his support for open models, an hour ago. reddit.com · 3d 20h ago
48 [BIG DATASET RELEASE] — SupraLabs/reasoning-corpus-4K-5M-v1 — Train your tiny SLMs to think! reddit.com · 3d 20h ago
46 As of JDK 27, Oracle engineers will thus stop maintaining the macOS/x64 port openjdk.org · 3d 19h ago
33 Show HN: Argus – An AI QA engineer: give it a URL and it tests your app argustest.com · 3d 19h ago
32 The world model remembers, the actor forgets: dissecting AI forgetting on 1 GPU arxiv.org · 3d 20h ago
29 Show HN: Compile computation graphs into transformer weights – no training github.com · 3d 20h ago
28 Justice Department Lawyers Reportedly Afraid to Put Things in Writing newrepublic.com · 3d 19h ago
27 The Internet of Snails: Escargotic commotion and the wood-wide web (2015) cabinetmagazine.org · 3d 19h ago
26 Ask HN: Why call it “inference” instead of just “model hosting”? news.ycombinator.com · 3d 18h ago
25 Jacobian Conjecture Refutation Reveals a Structural Limit of AI Interpretability ctolunchnyc.substack.com · 3d 19h ago
25 Show HN: Drive your real logged-in Chrome from Claude Code and Codex (MCP) github.com · 3d 19h ago
23 Tell HN: ChatGPT exports do not contain all conversation messages news.ycombinator.com · 3d 19h ago