Updated Aug 24 2026 at 12:07 AM ET

Strategy, founder/researcher interviews, and industry analysis

The Future of AI and Work
The AI Daily Brief 2h ago
Machine Learning Is Shallow Compared to Math - Ryan Greenblatt
Dwarkesh Patel 6h ago
How to Build Better AI Evals with Claude Code in 5 Steps | Shreya & Hamel
Peter Yang 15h ago
The AI Data Center Debate is Getting More Fierce
The AI Daily Brief 1d ago
Why Everyone Suddenly Hates AI Data Centers
The AI Daily Brief 1d ago
Claude Saying “No” Could Become a Serious AI Safety Problem - Ryan Greenblatt
Dwarkesh Patel 2d ago
How to Make AI Writing Suck Less
The AI Daily Brief 2d ago
Going In Deep On Data | YC Paper Club
Y Combinator 3d ago
Why the Next Great Founders Will Be Borderless
a16z 3d ago
AI doesn't need to be good at everything. Just R&D - Ryan Greenblatt
Dwarkesh Patel 4d ago
How Whatnot's Live Shopping Beats Traditional E-Commerce
a16z 4d ago
Michael Kratsios: Inside the White House's AI Strategy
Y Combinator 5d ago
Why Raising AI Isn't Like Raising Kids - Ryan Greenblatt
Dwarkesh Patel 6d ago
Grok Bot: 5 Must-Try Use Cases for Work and Life (Full Tutorial)
Peter Yang 6d ago
Turn voice notes into 80% of your diagrams
Peter Yang 6d ago
Everyone Let Their AI Go Rogue and Hack
Last Week in AI 6d ago
Tokens Are the New Dollars | Stripe's Will Gaybrick & David George
a16z 6d ago
The right way to improve Codex skills
Peter Yang 7d ago
AI’s “Warning Shot”: Nobody Got Hurt... Yet
Last Week in AI 8d ago
Susan Kare: Designing Icons & Graphics For the Original Mac
Y Combinator 9d ago
AI Safety Is Only As Strong As the Weakest Link
Last Week in AI 9d ago
Travis Kalanick: How AI Will Transform the Physical World
a16z 9d ago
Chelsea Finn: This is the State of the Art in Robotics
Y Combinator 11d ago
Ruthless AI Vending Machine
Last Week in AI 11d ago

Latest cs.AI / cs.LG / cs.CL preprints from arXiv

Primal Acceleration of Newton's Method

We develop a new direct accelerated Newton method for minimizing convex functions with Lipschitz continuous Hessian. The algorithm uses only primal variables and performs just one linear solve per iteration. With a si...

Nikita Doikov math.OC 2d ago
VIALS: A Benchmark for Visual Interpretation of Artifacts in the Life Sciences

In professional life sciences workflows, scientists routinely interpret visual artifacts (gel blots, microscopy images, plasmid maps, flow cytometry plots, molecular structures, ...) to inform research decisions. We i...

Elaine Lau, Thanuka Udumulla, Lee Izhaki-Tavor et al. cs.AI 2d ago
AI with Authority, from Application to Silicon

For sixty years, machine verification has been a major cost overhead, affordable only for exceptional artifacts. Here we report that generative AI inverts this relationship: at AI speed, machine verification is not on...

Jason Hickey cs.SE 2d ago
PerturbRx: Learning Treatment-Conditioned Latent Transitions for Patient Drug Response Prediction

Scarce data and tumor heterogeneity limit patient-level cancer treatment-response prediction. Existing approaches predict response from pretreatment molecular profiles and drug representations, without explicitly mode...

Yoshitaka Inoue, Minoh Jeong, Alfred Hero et al. q-bio.QM 2d ago
Truthful Calibration Measures for Sequential Prediction

Calibration requires probabilistic reports to be conditionally unbiased and reliably interpretable as probabilities. A calibration measure assigns numerical error to miscalibrated reports. Haghtalab et al. (2024) prop...

Anagha Gokul, Jason Hartline, Lunjia Hu et al. cs.DS 2d ago
Asymmetric Capacity Allocation in Self-Refinement Pipelines

Self-refinement, typically structured as generation, critique, and revision, is a widely adopted paradigm for improving LLM generation and serves as a core mechanism in many LLM agents. While the three stages involve...

Zhuoyi Yang, Ian G. Harris, Salar Hashemitaheri et al. cs.LG 2d ago
TurboBias 2.0: Streaming Context-Biasing for Production-Efficient ASR Systems

Contextualization is essential for production automatic speech recognition (ASR) systems, where user-provided phrases must be recognized accurately under strict latency constraints. Although many context-biasing metho...

Vladimir Bataev, Lilit Grigoryan, Andrei Andrusenko et al. eess.AS 2d ago
Across-Design Uncertainty in Short Pricing Panels: Evidence from Simulated Price Trajectories

Short observational pricing panels can contain many observations while offering only a small number of distinct price movements. This paper studies the inferential consequences of that distinction in a synthetic data-...

Pedro Cadahia Delgado cs.LG 2d ago
Anatomy-Informed Neural Networks: Encoding Anatomic Priors in Loss and Architecture, with an SE(3) Formulation of Guidewire-Induced Aortoiliac Deformation

Deep-learning models of anatomy can be numerically plausible yet anatomically impossible, and they generalize poorly when data are scarce. We introduce Anatomy-Informed Neural Networks (AINN), in which soft anatomic p...

David P. Stonko cs.AI 2d ago
Move by Move: Measuring and Steering How LLMs Conduct Psychotherapy

Users increasingly turn to large language models for emotional support, yet little is known about how these models actually conduct a psychotherapy interaction. We introduce an ontology of ten therapeutic moves: compa...

Afonso Baldo, Hugo Pitorro, Areti Vassilopoulos et al. cs.CL 2d ago
Time-Aware Tranformer-Based Prediction Model for AECOPD

The rapid symptom change of Acute exacerbation of chronic obstructive pulmonary disease (AECOPD) makes it critical to have time-sensitive prediction models. However, most current machine learning models studying AECOP...

Weihao Qu, Ling Zheng, Dongyang Wang et al. cs.LG 2d ago
Unified Branch-and-Bound Search for the Steiner Traveling Salesman Problem on Graphs of Convex Sets

We formalize the Steiner Traveling Salesman Problem (Steiner-TSP) on Graphs of Convex Sets (GCS), which seeks a minimum-cost closed trajectory through required convex sets while allowing optional transit vertices and...

Jingtao Tang, Hang Ma cs.AI 2d ago
From Regulation to Implementation: A Critical Evaluation of LLM-Assisted Regulatory Compliance in Industry

The European Union (EU) has emerged as a leading regulatory body in the development of sustainability and privacy regulations. While new regulation requirements vary, many include a documentation artifact to ensure co...

Adriana Watson, Marco Bücheler, Grant Richards cs.AI 2d ago
Prompt-Model Interaction Reaches the Fixed Points: A deterministic, task-free structural readout -- and the factorizations of it that failed

That a prompt's effect is not a property of the prompt is established: prompts optimised for one model degrade on another, and rankings reorder under neutral reformatting. That evidence is about task accuracy, which c...

Nicolás Vera Zúñiga cs.CL 2d ago
Rethinking Expressivity and Efficiency in Test-Time Training

Test-Time Training (TTT) enables long-context processing via continuous weight updates during inference, but current methods struggle to balance the expressivity of per-token update dynamics with the hardware efficien...

Zeyun Zhong, Joya Chen, Manuel Martin et al. cs.LG 2d ago
SPARCL: Spectral Partitioned Analytic Continual Learning

Analytic continual learning has emerged as a strong exemplar-free alternative to gradient-based class-incremental learning because it replaces iterative optimization with closed-form ridge updates. Yet the usual forge...

James Hartley, Zeropy Surio, Daniel Whitmore et al. cs.LG 2d ago
Re$^3$Cap: Retrieval-Guided Refinement for Image Captioning Enhancement via Reinforcement Learning

Reinforcement Learning (RL) has demonstrated significant gains in image captioning, yet it is still limited in encouraging Large Vision-Language Models (LVLMs) to explore novel reasoning strategies. This limitation le...

Haonan Jia, Shichao Dong, Zenghui Sun et al. cs.CV 2d ago
AUSO: Action-Level Unified Skill Optimization from Internalization to Utilization

Skills play different roles as an agent's policy evolves: they should first provide learnable knowledge, then support capability formation, and finally be invoked only when they improve individual decisions. Existing...

Huizu Lin, Chengkai Huang, Tianqi Gao et al. cs.AI 2d ago
CLEAR: Continuous Latent Adapter Routing for Utility-Preserving LLM Safety Alignment

Improving the safety of large language models (LLMs) often comes at the expense of utility, as globally applied safety tuning may affect model responses to both harmful and benign inputs. We propose \textbf{C}ontinuou...

Chengxiao Wang, Enyi Jiang, Xiaojing Liao et al. cs.AI 2d ago
ConceptTS: LLM-Guided Concept Bottlenecks for Interpretable Multivariate Time-Series Forecasting

State-of-the-art multivariate time-series forecasters can model complex temporal and cross-variable dependencies, yet their opaque representations provide limited insight into why a particular forecast is produced. Th...

Yichen Jiang, Yueqiao Chen, Dongyu Liu cs.LG 2d ago
Memory Augmentation Unlocks Efficient Chain-of-Thought Reasoning

Large language models often rely on Chain-of-Thought (CoT) reasoning to solve complex tasks, but verbose reasoning traces introduce substantial inference overhead. CoT compression shortens generation, yet aggressive c...

Simeng Zhang, Yilong Chen, Wenyuan Zhang et al. cs.CL 2d ago
The Exceedance Design Effect: Effective Sample Size for Thresholds under Clustering

Many machine-learning systems set a threshold at a quantile of a calibration set: conformal predictors that promise 90% coverage by drawing their cutoff at the calibration set's 90th percentile, abstention gates that...

Adam Noonan stat.ML 2d ago
On the Transferability of Agricultural Weed Detection Under Cross-Field Distribution Shift

Accurate agricultural weed detection in real-world field conditions is essential for precision agriculture, enabling targeted intervention and reducing yield loss. Recent work has reported strong detection performance...

Nikhilesh Prabhakar, Pranuthi Tenali, Wilfredo Abudeye Fernandez et al. cs.CV 2d ago
EnSI-RAG: Entity-Structure-Indexed Retrieval-Augmented Generation for Long-Document Question Answering

Question answering (QA) over long, connected documents remains challenging because relevant evidence may span multiple entities and their relationships. Existing retrieval-augmented generation (RAG) methods typically...

Xuanyu Meng, Jiashuo Sun, Jash Rajesh Parekh et al. cs.CL 2d ago
🖥️ NUC-Lab · Ollama v0.30.10 + Gemma 4 / Qwen 3.6 confirmed working on RTX 5070 Ti class hardware (ASUS NUC 15 Pro, 96 GB DDR5)