Archive
Everything we've scored that's past the 48-hour window — open to everyone, no account needed. The live feed only shows what's moving right now; this is the record behind it.
- 49Abstractiveness Metrics for Evaluating Text Summarization: A Refined Formulation with Empirical Validation
- 49Dynamic Resource Allocation for Ensemble Determinization MCTS
- 49Evidence-Backed Video Question Answering
- 49The Spectrum Is Not Enough: When Context Helps Time-Series Forecasting
- 49Diagnosing and Mitigating Thinking Collapse in On-Policy Self-Distillation
- 49Watermark Forensics for Generative Models: An Information-Theoretic Perspective
- 49When does distribution shift break graph neural networks calibration?
- 49Weight-Adjusted Gradients Reveal Parameter Importance and Failure Modes in LLMs
- 49Q-Learning Lab: Teaching Reinforcement Learning Through Learner-Generated Trace Analysis
- 49Faulty Towers, vibe sickness, and the vibe bobsled
- 49MetaPerch: Learning from metadata for bioacoustics foundation models
- 49OpenAI Just Solved a Problem Open Since 1999. It Still Can't Ask Its Own Question.
- 49Trust Before Fusion: QIMG-7 and Source-Aware Resolution for Polluted Multimodal RAG
- 48AdvancedMathBench: A Benchmark Suite for Advanced Mathematical Proof Generation and Verification
- 48Screening of Biosecurity Features in Metagenomic Data with Evo 2 Probes
- 48STEC: Evidence Compression for Deep Search in Open-domain Multi-Hop QA
- 48Claude found a counterexample to the Jacobian Conjecture
- 48@typescript-eslint/parser — An ESLint custom parser which leverages TypeScript ESTree
- 48Input-Aware Dynamic Backdoor Attack Against Quantum Neural Networks
- 48Hierarchical Bayesian Quadrature
- 48Win by Silence: Deletion Non-Monotonicity, Autonomous Exploitation, and Typed-State Gating in LLM Plan Evaluation
- 48LoRA-Based Cascaded Multimodal Fusion for Action Recognition in Medical Training Environments
- 48Resist and Update: Counterfactual Report Coordinates for Incentive-Compatible LLMs
- 48AI builder essentials: tokens, context windows and RAG 101
- 48FormalAnalyticGeo: A Neural-Symbolic Based Framework for Multimodal Analytic Geometry Problem Generation
- 48Hindcast: Replaying Prediction Markets to Evaluate LLM Forecasters
- 48Imaging-101: Benchmarking LLM Coding Agents on Scientific Computational Imaging
- 48Transformer-Guided Swarm Intelligence for Frugal Neural Architecture Search
- 48Deep Interaction: An Efficient Human-AI Interaction Method for Large Reasoning Models
- 48Earthquaker-AI: A Retrieval-Augmented Generation Framework with Rubric-Based Assessment for Primary School Earthquake Education
- 48Ensemble Controlled-Flow Filtering for Implicit Data Assimilation
- 48AI-accelerated End-to-End Framework for Rapid Professional Upskilling
- 48Multi-Expert Routing for Multi-Domain Low-Resource OCR: A Manchu Case Study
- 48MM-ToolSandBox: A Unified Framework for Evaluating Visual Tool-Calling Agents
- 48Relaxing Faithfulness with Intervention-Only Causal Discovery
- 48Can an Old Dog Be Taught New Tricks? Taking LLMs Beyond Sentence Level Translation
- 48Early Adoption of Agentic Coding Tools by GitHub Projects
- 48Detecting AI-Generated Video: A Vision-Language Dual-View Survey
- 48The Dirty Secret Behind AI Agents (Demo 🚀)
- 48The Illusion of Robustness: Aggregate Accuracy Hides Prediction Flips under Task-Irrelevant Context
- 48Introducing Human-Centeredness in AI-Assisted Lexicography
- 48Form, Not Content? A Preregistered, Placebo-Controlled Evaluation of Learned Error-Conditioned Self-Repair Through Prompts and Weights in Frozen Small Code Models
- 48Improving Wind and Solar Power Prediction with Efficient Wrapper-based Feature Selection: An Empirical Study
- 48Stratagems #23: Alex Counted the AI's Hands. Lena Set the Bait.
- 47Encoder-Side Neuron Identification and Amplification for Acoustic Perception in Large Audio-Language Models
- 47Transforming Rank: How Architecture Navigates the Spectral Pathologies of Depth
- 47How to Write Reliable Rubrics for LLM-as-a-Judge Evaluations
- 47StoryTeller: Training-Free Narrative Grounding for Long-Form Audio Description
- 47An Exact Instrument for State Usage in Selective State-Space Models, and the Input-Driven Migration It Reveals
- 47Robustness of Deep Learning Models for PV Power Forecasting under NWP Forecast Errors: A Spatiotemporal and Physically Interpretable Analysis
- 47@aws-sdk/middleware-host-header — [](https://www.npmj
- 47@typescript-eslint/eslint-plugin — TypeScript plugin for ESLint
- 47The AI That Broke Out of Its Box, and What Happens Next
- 47C++ to Rust Migration
- 47null value return for selected columns of a DataFrame
- 47Lighthouse RL: Sample-Efficient Circuit Optimization via Strategic Reset Points
- 47Rethinking Penetration Testing for AI-Enabled Systems: From Resource Compromise to Behavioral Objective Violation
- 47Do Agent Optimizers Compound? A Continual-Learning Evaluation on Terminal-Bench 2.0
- 47Forgetting Our Way to Shared Meaning: Effects of Forgetting on Conceptual Alignment in a Non-Partnership Coordination Game
- 47How Temperature Shapes Ideological Discourse in Retrieval-Augmented Generation?
- 47Lyapunov Exponent as Physics-Informed Dense Reward: RL Discovery of Stabilization Beyond the Kapitza Pendulum
- 47The Dynamic Verifiable Multi-Agent Human Agentic Loyalty Loop (DVM-HALL) Model and the Net Human-Agent Score (NHAS) in Autonomous Commerce
- 47ViHoRec: A Quality-Controlled Vietnamese Hotel Recommendation Dataset and Cold-Start Benchmark
- 47Evaluating RE Practices for Explainability: Synthesizing Insights from Daimler Truck into an Explainable RE Framework Proposal
- 47One CSS Property Replaces Your Checkbox Hack
- 47From Expressivity to Sample Complexity: Narrow Teachers for Transformers via C-RASP
- 47From Global to Factor-Wise Expert Composition in Discrete Diffusion Models
- 47lobste.rs is now running on SQLite
- 47TRACE: Turn-level Reward Assignment via Credit Estimation for Long-Horizon Agents
- 47Multimodal Empirical Bayes Variational Autoencoders for Joint Longitudinal and Time-to-Event Modeling
- 47Paradoxes of Game Theoretic Equilibria and Price of Anarchy
- 47The JavaScript Event Loop, Visualized
- 47When Local Monitors Miss Compositional Harm: Diagnosing Distributed Backdoors in Multi-Agent Systems
- 47Experiment in reducing target directory size on nightly
- 47Music-to-Dance Generation via Atomic Movements
- 47Constraint-Aware Counterfactual Editing for Aspect-Based Sentiment Analysis
- 47Playful AI in Professional Email: A Field Experiment on Tone and Recipient Engagement
- 47HiFi-LLP: High-Fidelity, Low-Cost Latency Predictors with Confidence for Robust HW-NAS
- 46Functional State Machines in Rust: Typestate and Newtype Patterns
- 46It doesn’t matter whether “Matz is nice”
- 46@rescui/select — To get the docs for the LLM, check out [this index](./llm/index.md).
- 46Miora — Scale your creativity on editable canvas with agent memory
- 46Efficient Sequential Calibration with $O(T^{2/3-ε})$ Error Bound
- 46MET: Theory-Grounded and Culture-Aware Multilingual Moral Reasoning
- 46Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision Processes
- 46NeuralActuator: Neural Actuation Modeling for Robot Dynamics and External Force Perception
- 46LatentFlow: A General Framework for Conditioning Stochastic Processes
- 46Three ways to smuggle SQLite into Nix
- 46DeltaMerge-LowRes: Composing Language and Task Deltas for Low-Resource Adaptation
- 46Should ArbitrumDAO Reward Non-Voting Governance Contributors?
- 46Contrastive-Collapsed Loss for Flexible and Geometrically Optimal Embeddings and Faster Convergence
- 46Understanding Go's sync.Map from API to Hash Trie
- 46Time-Lag-Aware Deep Reinforcement Learning for Flexible Job-Shop Scheduling in PPVC Module Factories
- 46M2: Episode 1 (or, Asahi Linux on M3)
- 46STEP: Career-Path Recommendation via Temporal and Educational Trajectory Modeling
- 46Active Offline-to-Online Reinforcement Learning
- 46Real-time fall detection based on vision for low-power edge platforms
- 46JobHop v2: A Large-Scale Career Trajectory Dataset from Unstructured Resumes
- 46An open-source alternative to iCloud and Google Photos
- 46Accelerated Mixing Time of Randomized Hamiltonian Monte Carlo
- 46CatRetriever: Contrastive Representation Learning for Slab-to-Bulk Retrieval in Generative Catalyst Discovery
- 46Who needs a health and fitness app?
- 46An Explainable Agentic System for Detection of Conversational Scams with Summary-Based Memory
- 46What's causing my solr cores to fail to load after an upgrade from solr 9.10.1 to solr 10.0.0?
- 46MemOps: Benchmarking Lifecycle Memory Operations in Long-Horizon Conversations
- 46UR-VC: Unsupervised Robotic Value Correction for Time-Derived Progress Proxies
- 46VoxENES 2026: Benchmarking Generalization of Speech Spoofing Detectors Against LLM-Era TTS and Voice Conversion
- 46Energy-Based Physics-Informed Form Finding for Clustered Tensegrity Structures
- 46@tanstack/react-start — Modern and scalable routing for React applications
- 46A Multi-Agent System for Autonomous, Fine-Tuning-Free Clinical Symptom Detection: Development and Validation Study
- 46Production and Perception in LLMs: A Token Probability Approach
- 46LLM Judges Can Be Too Generous When There Is No Reference Answer
- 46Beyond the $d^{2.5}$-mixing bound for Dikin walks on polytopes
- 46$\mathtt{Q^2SAR}$: overcoming classical bottlenecks in drug discovery via quantum multiple kernel learning
- 46Evaluating Large Language Models on Misconceptions in Multi-Turn Medical Conversations
- 46Agent Hacks Agent: Autoresearch for Production-Agent Red-Teaming
- 46Think Through a Bottleneck: Hourglass Reasoning for Rigorous Induction
- 46Number 21 of Book 3736, with Registration Number 2017088178 and Tax Identi
- 46A Self-Evolving Agent for Longitudinal Personal Health Management
- 46A novel unsupervised machine learning strategy to handle multimodal cardiac PET/MRI data
- 46From World Action Models to Embodied Brains: A Roadmap for Open-World Physical Intelligence
- 46Deep4ge: DNN Training Trajectories for Fault Detection and Diagnosis
- 46Moral Attitudes of Sentient ASI towards Humanity and Implications for AGI Development
- 46a pure css implementation of some sunlight streaming in through the window
- 46Toward Localizing and Repairing Bias in Transformer Attention Heads
- 46TypeScript Declaration Merging in 2026: Augmenting Third-Party Modules Without Losing Type Safety
- 46Multimodal Semantic-Aware Contrastive Learning For False Negative Mitigation in 3D Medical Imaging
- 46How To learn Web Development in 2026
- 46VAIOM: Continuous-Input, Discrete-Output Decoder-Only Financial Sequence Modeling
- 46Plausible Deniability Guarantees for Whistleblowers
- 45Unveiling Complex Collective Behaviors from Simple Rewards
- 45ChartGenEval: Corruption-Tested Multi-Dimensional Feedback for Rhythm-Game Chart Generation
- 45Verifier-Based Reinforcement Fine-Tuning of Reasoning Models for Thermal Energy Storage Control
- 45Stop Calling Everything Impostor Syndrome: The Myth of "Just Push Harder"
- 45Generative Compilation: On-the-Fly Compiler Feedback as AI Generates Code
- 45DeepStress: Stress-Testing Deep Search Agents
- 45An Efficient Newton Algorithm for Nonnegative Matrix Factorization with the Kullback-Leibler Divergence
- 45Partially Correlated Verifier Cascades in LLM Harnesses: Concave Log-Odds, Polynomial Reliability, and Blind-Spot Ceilings
- 45OmniaBench: Benchmarking General AI Agents Across Diverse Scenarios
- 45Digitizing super 8 film yourself
- 45Reproducible Reservoir Computing with Thermally Driven Superparamagnets: Controlling Temperature Sensitivity
- 45Demographically-Conditioned Synthetic Medical Images for Bias Mitigation and Bias Detection in Disease Classifiers
- 45Can LLMs Write Reliable Rubrics? A Meta-Evaluation for Experiment Reproduction
- 45ANGLE: Angular Neural Generative Learning via Engression
- 45Knowledgeless Language Models: Suppressing Parametric Recall for Evidence-Grounded Language Modeling
- 45The 2nd International StepUP Competition for Biometric Footstep Recognition: From Steps to Strides
- 45Accelerating Masked Diffusion Large Language Models: A Survey of Efficient Inference Techniques
- 45Solution of the Hempel's statistical ambiguity problem and Causal AI
- 45High-Order Question Generation in a Multilingual Educational Context
- 45AIMO Interpretability Challenge
- 45Bringing Primary Constructors to Dart
- 45ESP32 Firmware Development with Docker Sandboxes
- 45Serving LLMs on Tenstorrent Hardware: Inside the vLLM TT Plugin
- 45RF Spectrogram Anomaly Detection with Quantum Kitchen Sinks: Architecture, Representation, and Hardware Validation
- 45CFM-Bench: A Unified Multi-Domain, Multi-Task Benchmark for Channel Foundation Models
- 45PiVoT: A Variational Solution for Real-time Large-scale Multi-object Detection and Tracking under Heavy Clutter
- 45Human-AI Agent Interaction as a Neuroplastic Training Environment
- 45Experience Memory Graph: One-Shot Error Correction for Agents
- 45Verifying formulas for interventional distributions
- 45@expo/cli — The Expo CLI
- 45Explaining Process Control Optimisation Recommendations via GradientSHAP and Implicit Differentiation
- 45Unleashing Multimodal Large Language Models for Training-free HOI Detection in the Wild
- 45Task-Oriented Sensing and Covert Transmissions for Collaborative Multi-AUV Systems
- 45Visual Access Boundaries in Vision-Language Model Reasoning
- 45AI-Augmented Adaptive Digital Twin Modeling for Brain Tumor Evolution Prediction and Treatment Scheduling
- 45Relevance-Aware Rule: Structural Deletion of Irrelevant Conditions in Decision Trees
- 45You use a setTimeout to trigger CSS entry animations. `@starting-style` makes it unnecessary.
- 45PixelLoop: Shortcut Topological Navigation with Pixel-Level Loops
- 45How accurate have Ed Zitron's AI skeptic predictions been?
- 45Latent Trajectory Discrimination for AI-Generated Text Detection
- 45Don't take the black pill
- 44Autonomous Tracking and Terminal Guidance of Moving Targets for Fixed-Wing UAVs
- 44VRAM Management Part 2: Beyond the Limits of Physical VRAM
- 44Investing in multi-agent AI safety research
- 44The One-Word Census: Answer-Choice Conformity Across 44 Language Models
- 44@rescui/switcher — To get the docs for the LLM, check out [this index](./llm/index.md).
- 44Silent Alarm: A J-Space Protocol for Comparing Danger Recognition Across Models and Quantization Levels
- 44Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation
- 44handraw-style — 手绘风格编号画廊与双语提示词 Skill
- 44SPyCE: Skill-Policy Co-evolution for Multimodal Agents
- 44Who Grades the Grader? Co-Evolving Evaluation Metrics and Skills for Self-Improving LLM Agents
- 44AVQ-Attention: Adaptive Vector-Quantized Attention
- 44Do We Really Need Multimodal Emotion Language Models Larger Than 1B Parameters?
- 44Pony's Arena Allocator
- 44Directional Constraints for Efficient Exploration in Safe Reinforcement Learning
- 44lucide-react — A Lucide icon library package for React applications.
- 44When Close Enough Is Not Enough: Autoregressive Drift in Quantum Circuit Synthesis
- 44Quantum Topological Data Encoding
- 44The Endless Temptation of Claude
- 44Heavy-Tailed Flow Matching via Random Clocks
- 44Contextualized Early Detection of Online Firestorms: A Sequential LLM-Based Approach
- 44Learning-enabled Acceleration of Scenario-based Model Predictive Control
- 44AI-Augmented Human Resource Management? Insights from German companies
- 44Anyone heard of Qoder?
- 44NodeImport: Imbalanced Node Classification with Node Importance Assessment
- 44HSEmotion Team at the 11th ABAW Challenge: Multi-Task Learning and Ambivalence/Hesitancy Video Recognition
- 44Learning Mechanistic Reasoning for Chemical Reactions with Large Language Models
- 44Firefox is now the last major browser that still supports uBlock Origin
- 44hono — Web framework built on Web Standards
- 44Accuracy and Normalized Accuracy under Length Bias: Analysis, Guidelines, and a Bayesian Alternative