100 Adversarial Style Optimization: Enhancing VLM Jailbreaks by GRPO-based Stylistic Triggers Optimization arxiv.org · 2m ago
100 Humanly: A Configurable and Traceable Environment for Human-AI Collaborative Writing arxiv.org · 2m ago
100 Leveraging External Knowledge for Historical Document Restoration via Retrieval-Augmented Large Language Models arxiv.org · 2m ago
100 Ground Truth First: A Longitudinal Evaluation Instrument for Agent Memory, and the Tenure Crossover in Memory-Architecture Rankings arxiv.org · 2m ago
100 Analysing Self-Harm Representations in Language Models: a Cross-Architecture Study arxiv.org · 2m ago
100 Enough is as good as a feast: A Comprehensive Analysis of How Reinforcement Learning Mitigates Task Conflicts in LLMs arxiv.org · 2m ago
100 Developing and Validating the Spanish Version of the Large Language Models Dependency Scale (LLM-D12-SP) arxiv.org · 2m ago
100 From Isolated Tasks to Structured Capabilities: A Multilayer Taxonomy for Large Language Models arxiv.org · 2m ago
100 Why Large Language Models and Humans Converge and Diverge in Evaluating Creativity arxiv.org · 2m ago
100 Skill Self-Play: Pushing the Frontier of LLM Capability with Co-Evolving Skills arxiv.org · 2m ago
100 From Obligation to Specification: A Survey on Validating EU AI Act Requirements in RE arxiv.org · 2m ago
100 Adaptive Depth in Looped Transformers: Diagnosing Learned Halting Gates and Trajectory Readouts arxiv.org · 3d ago
100 From Atoms to Entropy: Optimal Noise Allocation for Diffusion Training in the Convex Regime arxiv.org · 3d ago
100 HypNO: A Graph-Based Neural Operator with Physics-Informed Message Passing for Hyperbolic Conservation Laws arxiv.org · 3d ago
100 Conflict Resolution under Degraded Surveillance in Air Corridors Using Multi-Agent Reinforcement Learning arxiv.org · 3d ago
100 SevDiff: Severity-Conditioned Diffusion for Long-Tail Conflict Trajectory Generation arxiv.org · 3d ago
100 Beyond SBDD: Geometric Deep Learning in Polypharmacology and Multi-target Drug Design arxiv.org · 3d ago
100 SenCos-GEM: SENet-Calibrated and Law-of-Cosines-Constrained Geometry-Enhanced Molecular Representation for Property Prediction arxiv.org · 3d ago
100 ClickGuard: Detecting and Spoiling Clickbait News with Informativeness Measures and Large Language Models arxiv.org · 3d ago
100 Stochastic Sampling is Epistemically Shallow: The Dimensionality Gap Between Temperature Variation and Model Diversity in LLMs arxiv.org · 3d ago
100 InferenceBench: A Benchmark for Open-Ended LLM Inference Optimization by AI Agents arxiv.org · 3d ago
98 Partnerships can keep open source sustainable stackoverflow.blog · 3d ago
43 BTL-3: A 27B open-weight agent model for agentic coding and structural tool use huggingface.co · 2d 23h ago
41 Amazon and Microsoft pivot cloud gaming strategies to target different players cnbc.com · 2d 22h ago
41 Show HN: Continuum – switch AI coding agents without re-explaining your project github.com · 2d 22h ago
41 Japanese AI Robots Used to Replicate Skilled Confectioners’ Abilities japannews.yomiuri.co.jp · 2d 23h ago
37 HN: FlowShelf –free,native macOS shelf/clipboard/screenshot app,on-device AI flowshelf.app · 2d 22h ago
33 Codex Slides: open-source AI slide studio powered by Codex. Prompt, repo to deck github.com · 2d 23h ago
1 Розборки ШІ : OpenAI зламала американську ШІ — платформу на порятунок прийшла китайська ШІ — модель itc.ua · 2d 23h ago