Archive
Everything we've scored that's past the 48-hour window — open to everyone, no account needed. The live feed only shows what's moving right now; this is the record behind it.
- 9Xiaomi-Robotics-1
- 9ParVL: Parallel Scaling and Expandable Compute Allocation for Multimodal LLMs
- 9SocietyBench: Forecasting Counterfactual Social-World Evolution
- 9WorldCup Arena: Prospective, Leakage-Free Evaluation of Frontier LLMs on a Live Tournament
- 9TurnSight: Turn-Level Hindsight Self-Distillation for Tool-Integrated Reasoning
- 9PAST-Bench: Benchmarking the Foundations of Recursive Self-Improvement in Personal Agents
- 9Test-Time Scaling in Reasoning LLMs: Inference Regimes, Evaluation, and Reproducibility
- 9Agogic: Performance-Timed Music Tokens for LLM-Native Text-to-Symbolic-Music Generation
- 9Adaptive deep nonparametric regression from dependent data under covariate shift
- 9Show HN: Bor – Open-source policy management for Linux desktops
- 9Manifold-Constrained Hyper-Connections for Parameter-Efficient Finetuning
- 9When Attention Goes Blind: Numerical Failure in ALiBi Positional Encodings
- 9Sequential Learner Modeling Using Multi-Relational Graph Convolutional Networks
- 9Reinforcement Learning for Code Optimization
- 9Assessment of Conditional Diffusion Model for Synthetic Histopathology Image Generation
- 9microsoft/Mage-Flow (text-to-image)
- 9Classical Hardware Acceleration of Quantum Autoencoders for Real-Time Anomaly Detection in Collider Experiments
- 9ClouDens: Operational Context-Aware Anomaly Detection for Large-scale Cloud System Monitoring
- 9Agentic Permissions Policy Algebra for Taint Confinement in LLM Agents
- 9Sky sphere representation in language models
- 9The Blessing of Dimensionality: How Near-Orthogonality in High-Dimensional Spaces Explains Temporal Portability
- 9Don't Trust the Label: License Laundering in AI Supply Chains
- 9Quasi-SVD: Learning a Lie-constrained matrix factorisation for real-time imaging
- 9A Model for Imbalanced Label Aggregation: A Focus on Minority-Class Detection
- 9string2string Studio: An Interactive, In-Browser Platform for String-to-String Algorithms
- 9InferScale: GPU-Native KV Injection for Personalized LLM Serving
- 9Can Large Language Models Recover Semantic Optimization Opportunities That Compilers Miss?
- 9Inference-Time Steering for Cross-Lingual Factual Consistency in LLMs
- 9Video-DeepResearch: Towards the Next-Generation Multimodal Deepresearch Agent
- 9Knowledge-Guided Multimodal Reasoning over Interacting Streams for Video-Level Ambivalence and Hesitancy Recognition
- 9alibaba-pai/MiniMax-H3-Fun-Controlnet-Union (text-to-video)
- 9COVAriance-Induced Fairness Gap Penalty for Subgroup-Fair Clustering
- 9Detecting Knowledge Inconsistencies Across Text, Tables, and Knowledge Graphs
- 9Thermodynamics-Informed Input Reparameterization for Neural Prediction of Real-Fluid Thermodynamic Properties in Supercritical Combustion
- 9Large Language Model for Operations Research Formulation Selection in Multi-Warehouse Inventory Allocation
- 9EIP-8369: VOPS Profiles for FOCIL Eligibility
- 9SGA: Plug&Play Geometric Verification for Educational Video Synthesis
- 9Polistemics: Evaluating LLMs as Information Mediators in Politics & Elections
- 9Attribution and Uncertainty Behavior of Learned Residual Gyro Correction for Gyro-Stellar Estimation
- 9How Does Alignment Tuning Shape Representations of Sycophancy and Related Cue-Induced Biases in LLMs?
- 9ReflectRL: Learning from Golden Negative Trajectories via Reflective-to-Direct Reasoning
- 9LLM-as-a-Coach: Experiential Learning for Non-Verifiable Tasks
- 9Should We Type or Talk to LLM Agents? A Comprehensive Study of Voice and Keyboard Input Perturbations
- 9DBMol: Design of High-Affinity, Target-Specific Small Molecules through Structure Prediction Models
- 9SciFigQual-Bench: A Benchmark for Scientific Figure Quality Assessment with Full-Manuscript Context
- 9Courteous Anticipation: Improving Long-Lived Task Planning in Persistent Shared Environments
- 9MeetingToM: Evaluating Multimodal LLMs on Theory-of-Mind Reasoning in Multi-Party Meetings
- 9Looping Is Not Reliability: State-Bound Evidence and Typed Revision Contracts for Agentic Code Repair
- 9Evaluating the Impact of Explainable AI on Trust in AI-Assisted Code Review
- 9In-Context Time Series Classification with Random Convolutional Features
- 9MODUS: Decoder-Only Any-to-Any Modeling of Diverse Modalities
- 9Information-Geometric Forward Policy Training in GFlowNets
- 9S3: Stable Subgoal Selection by Constraining Uncertainty of Coarse Dynamics in Hierarchical Reinforcement Learning
- 9HalluTruthQA-4K: A Fine-Grained Corpus and Annotation Process for Arabic Hallucination Detection and Truth Verification
- 9Selection Shapes the Boundary: A Preregistered Replication of Monotonicity and Label Agreement in Unselected NLI Populations
- 9Empowering On-Device Model Adaptation with an Edge AI Inference Accelerator
- 9LoRA Speedrun – a public wall-clock leaderboard for fine-tuning techniques
- 9Can We Break LLMs Out of Self-Loops? Fine-Grained Reasoning Control with Activation Steering
- 9Sound Probabilistic Safety Bounds for Large Language Models
- 9VDAR-Router: Adaptive LLMs Routing via Verbalized Query Difficulty Analysis Retrieval
- 9On the post-hoc Evaluation of PDE Discovery: A Multifaceted Challenge of Scientific Advancement
- 9PIVOT: Efficient Query-Group Indexing for Token-Level Sparse Attention
- 9Separating quantum circuits from classical LLMs
- 9The Price of Reasoning: Cost-Quality Tradeoffs in Reinforcement Learning for Neural Machine Translation
- 9Interpretable Adaptive Sampling for LLM Test-Time Scaling
- 9Interpretable Fuzzy Rule-Based Regression Extension for Ex-Fuzzy Library
- 9SciForma: Structure-Faithful Generation of Scientific Diagrams
- 9Self-supervision drives representational convergence in medical foundation models more than clinical supervision
- 9Artificial Intelligence and Innovation Ecosystem: Evolutionary Developments, Challenges, and Future Directions
- 9The Label Complexity of Class-Conditional Coverage under Distribution Shift
- 9A game theory for foundation models shows new paths to rational cooperation through similarity inference
- 9SIREN: Towards End-to-End Extreme-Weather Early Warning with Experience-Grounded LLM Agents
- 9Judge-dependent safety gains and model-specific helpfulness costs of evidence-sufficiency prompting in clinical LLMs
- 9Soft-Constrained Optimization of Latent Space in Variational Autoencoders
- 9AdaFlash: Adaptive Speculative Decoding via On-Policy Distilled Diffusion Drafters
- 9D-Score: A Spectral Hidden-State Signal for Hallucination Detection in Large Language Models
- 9WorldCupArena: Fine-Grained Evaluation of Language Models and Deep-Research Agents on Football Forecasting
- 9Which Values Do LLMs Confuse? A Schwartz-Based Recognition Study
- 9From Data to Device: ELMOD An Efficient German-First 2.7B Language Model for Mobile Inference
- 9PYPM-GGD: Pitman-Yor Process Mixture with Generalized Gaussian Density using ADAM
- 9PoTRE: Test-Time Reasoning inspired by Cognitive Heterogeneity
- 9CADER: Confidence-Aware Dynamic Evidence Reasoning for Long-Video Understanding
- 9Enhancing Rubric-based RL via Self-Distillation
- 9The Maskability Index: Predicting Task-Objective Alignment in Pretrained Language Models
- 9Beyond Score Prediction: LLM-Based Essay Scoring and Feedback Generation via Reinforcement Learning with Rubric Rewards
- 9SelectInfer: Selective Neuron Loading and Computation for On-Device LLMs
- 9Sparse Evidence Can Suffice: Agentic Evidence Seeking for Multimodal Video Misinformation Detection
- 9Kimi K3 — The world's first open 3T-class model
- 9Evaluating Fuzz Testing for Reinforcement Learning Agents
- 9Generalised Bellman recurrence and three dualities in sequential decision-making
- 9Modeling turn-taking with distant viewing: investigating silence thresholds in human and AI-generated discourse
- 9TACT: Taxonomy-Aligned Post-Training for Pedagogically Adaptive English Tutoring
- 9Breaking the $T^{3/4}$ Barrier for Regret Minimization With Bi-Dimensional CDFs
- 9The Ethics of Autonomous AI Agents for Offensive Security
- 9Sobek: Streaming Equivariant Tensor Product Convolutions
- 9Computing on the Fly: Navigating a Vision for the Future of Drone Computing
- 9Pushing the Frontier of Full-Song Generation: Hierarchical Autoregressive Planning Meets Flow-Matching Rendering
- 9Exposure is Optional: Learning Unlike Coordination in Language Models
- 9SGN: A Similarity-based Generative Network for Data Generation under Distribution Shift
- 9Muon Meets Mamba: Spectral Optimization for State Space Models
- 9Assessment in Team Problem-Solving Exercises in Computing Education
- 9LLM-SoccerArena: Benchmarking LLMs on Real-World Predictions in Sports
- 9No link between acetaminophen use during pregnancy and adverse birth outcomes
- 9The Visual Bottleneck: Sparse-Frame Adaptation of MLLMs for Joint Spatial-Temporal Video Grounding
- 9The balance between compactness and forecast accuracy of data-driven latent-space reduced-order models in controlled wake flows
- 9Bit-Accurate FPGA Evaluation of Learned Feature Gating in a Fixed-Point Fourier-Feature Automatic Modulation Classifier
- 9DSCH-Loss: A Dynamic Semantic Channel Objective for Deep Semantic Hashing
- 9Hardware Mechanisms to Dynamically Throttle AI Performance
- 9Human Grounded Evaluation of Large Language Models for Optical Network Automation
- 9MIRA-Ev:A Benchmark for Granular Evidence Detection and Relational Reasoning in Clinical Exams
- 9Conservative Query and Adaptive Regularization for Offline RL Under Uncertainty Estimation
- 9TRACE-CTI: Auditable Post-Extraction Governance of TTP Claims with Knowledge Graphs
- 9Pancasila-Dilemmas: Evaluating Large Language Models on Indonesian Human Value Dilemmas Grounded in Pancasila
- 9How do I prevent redaction rectangles from removing surrounding text?
- 9Hierarchical Group-Conditional Conformal Risk Control for Selective Prediction in Language Models
- 9Logic Before Language: Pre-pretraining on Formal Derivations Fosters Skill Acquisition and Compressibility
- 9Autoresearch with Coding Agents: Generalizers and Metric-Maximizers on Quran Recitation Data
- 9ATLAS: A Foundation Neural Sampler for Amorphous Materials
- 9EgoPlay: Event-Triggered Video Editing for Egocentric Streams
- 9On the Systematic Challenges of Culturally Loaded Machine Translation: Dream of the Red Chamber as the Cultural Lens
- 9Free energy landscape of Dense Associative Memory
- 9Adaptive Adversaries: A Multi-Turn, Multi-LLM Benchmark for LLM Agent Security
- 9Latent Reward Registers for Diffusion Preference Alignment
- 9BettiSplit: Topology-Guided Privacy-Aware Split Learning Against Feature Inversion and Gradient Leakage
- 9Robust Low-Tubal-Rank Tensor Completion under Cross-Concentrated Sampling
- 9Adaptive Bayesian Online Learning via Expert Aggregation
- 9LOCKS: Page-Local Compact Key Summaries for Efficient Long-Context Decoding
- 9A Physics-Flavored Transformer Network for Parametrizing Contraction Dynamics of Engineered Skeletal Muscle Tissues
- 9EchoBridge: Long-Tail-Aware ECG-Echocardiography Text Alignment for Echocardiography-Derived Cardiac Findings
- 9PhaseAware: Interpretable Human-in-the-Loop Rehabilitation Scoring with Boundary Monitoring
- 9ABot-World-0: Infinite Interactive World Rollout on a Single Desktop GPU
- 9LLM-Assisted Ontology Engineering and Construction of a French Legal Knowledge Graph
- 9Dynamical and Optimization Trade-offs of Levi--Civita Coordinates for Learned Close-Encounter Dynamics
- 9An Early Warning of Emerging Biosecurity Risks in Frontier LLMs
- 9Zing: Social Mind for LLMs
- 9So Reddit has decided that plain HTML is unsafe
- 9Learn WebGPU for C++
- 9Agentic Real2Sim: Physics-based World Modeling with Vision-Language Agents
- 9The Transformer Revolution, Part 1: Dynamic Processing through Output- Weight Interconnections
- 9Equivariant Music Transformer
- 9From transcription to semantic corpus analysis: unsupervised learning of sentence representations for ancient languages
- 9When and Where to Look: Adaptive Visual Evidence Scheduling for Efficient Long Video Understanding
- 9PIER: Physics-Informed Environmental Retrieval for Time-Series Modeling
- 9Implementing Causal Perception: Competing SCMs and Situated Fairness
- 9Task-Conditional Faithfulness Auditing of Multimodal LLMs for Grid Diagnosis
- 9User-Centric Modeling of Transactional Sequences with Explainable State Space Models
- 9SEE: Structure-aware Exploring \& Exploiting for Long-horizon GUI Agent Trajectory Synthesis
- 9The K-SCAN Clustering Algorithm
- 9DQAOA-GPT: AI-Accelerated Distributed Quantum Optimization for Combinatorial Problems
- 9Reasoning Before Translation: Enhancing Legal Machine Translation with Structured Reasoning
- 9Automated Extraction of Techno-Economic Data from 76,000 Energy System Studies
- 9The Shared Discovery Paradox: How a One-Answer Rule Turns Better Information into Worse Search
- 9Is the postfix expression `f` sequenced before the function call expression `f()`?
- 9Adaptive Mamba Neural Operators
- 9Anticipate Before Acting: Future-State-Conditioned Vision-Language Navigation
- 9From Machine Learning to Large-Scale EO Products: Best Practices for Making Maps
- 9Neural Kolmogorov Equations: Parallelizable Learning of Stochastic Dynamics under General Noise
- 9HalluTruthQA: A Fine-Grained Benchmark for Hallucination Detection, Localization, and Explanation in Arabic Question Answering
- 9Source-Free Controlled Adaptation of Teachers for Continual Test-Time Adaptation
- 9TRUAV: Distributed Multi-Agent Reinforcement Learning for Trajectory Planning and Routing Enhancement in UAV-Aided IoT-Enabled VANETs
- 9Small, Free, and Effective: Orchestrating Open-Weight Small Language Models to Outperform Single LLM for Malware Analysis
- 9comfyui-minimax-h3-audio-T8 —
- 9AI Strategy: How to Choose What AI Product to Implement
- 9ELSAA: Efficient Low-Rank and Sparse Attention Approximation for Training Transformers
- 9An Exact Counterexample to Carlson's Associated-Prime Depth Conjecture from a Group of Order 128
- 9You're rethrowing errors and losing context. `Error.cause` fixes that.
- 9Outcome-Confounded Local Supervision in On-Policy Distillation
- 9Boundary-Adapted PINNs for Elliptic Dirichlet Problems: $H^2(Ω)$ A Priori Error Bounds with Application to Mean Escape Time Computation
- 9FlowCTS: On-policy Continuous Trajectory Supervision of Flow Models
- 9AdaHome: An Adaptive Smart Home Assistant using Local Small Language Models
- 9Stress-Testing EEG Foundation Models for Clinical Decoding: Dataset Identity and Targeted Negative Controls
- 9Natural Language Access to Domain-Specific Metadata: A Reusable Framework for LLM Query Generation
- 9surprisal is Not a Theory
- 9Low-Rank Dependence Decomposition via Accelerated Symmetric Non-negative Matrix Factorization
- 9Statistical Inference for Rank Allocation in Low-Rank Adaptation
- 9DecoupleMix: Decoupled Ratio Search and Convex Allocation for Scalable VLM Data Recipes
- 9Systematic Analysis of Large Language Models and Transformer-Based Machine Translation for English-Tamil and Tamil-English Across Diverse Datasets
- 9Physics Transformer: Tailoring Transformer for General PDE Prediction
- 9The Quadrilateral Loss: Additivity as a Measurable Behavior of Dense Neural Networks
- 9Breaking the Homogeneity Assumption: Specialized Multi-Generator Adversarial Learning for Rare Failure Detection in Predictive Maintenance
- 9Grok Build is open source
- 9Turns Out I Never Knew What Node.js Was
- 9DeepSeek-V4-Flash Update
- 9Show HN: Maple-Preview – Ternary 20B MoE running at 120 tok/s on a iPhone
- 9OLEDLM: A Unified Language Model for OLED Molecular Design
- 9On Optimization Complexity of Second-Order Certified Unlearning
- 9Hierarchical Soft Actor-Critic for Sparse-Reward Long-Horizon Reinforcement Learning
- 9What is the idiomatic way to track dependencies for asynchronous callbacks in Svelte 5?
- 9Incomplete Observations Boost Evolutionary Performance in Ocean Modeling
- 9E-Bench: Benchmarking Multi-Step Tool-Use Agents in Real-World Product Scenarios
- 9Distributional Split Criteria for Random Forests: Extensions, Shrinkage, and the Robustness of Mean Splitting
- 9StreamHOI: Interaction-aware Temporal Memory Adaptation for Streaming HOI Video Generation
- 9Instance Hardness-Based Relevance for Imbalanced Regression
- 9Hard Guarantees at a Measured Price: Entropy-Stable Learned Finite Volumes for Compressible Flow
- 9Emacs Is a Lispboard
- 9Audio-Zero: Label-Free Self-Evolution for Fine-Grained Audio Reasoning
- 9How to Remove EXIF and GPS Metadata Before Sharing a Photo
- 9Plausibility-Driven Prioritization of Candidate Biomedical Annotations
- 9Self-organizing Architecture of Receptron Units: a Hardware-Aware Framework for Edge Intelligence
- 9MIRAGE: Multi-scale Lesion-Informed Representation with Auxiliary Guidance for MRI Contrast Enhancement