Signals 4
· free daily AI digest
cs.AI — items we recorded
56 item(s) from this outlet appeared in our collection.
Regularized Emphatic Temporal-Difference Learning: Stability under Constant Stepsizes
2026-09-18
BioPhys-Bridge: A Benchmark for Interdisciplinary Scientific Reasoning in Physics-Grounded Biological Research
2026-09-18
What Do We Expect from LLMs? Mapping the Design of LLM Benchmarks
2026-09-18
Position: It is Time to Virtualize Foundation Models with a Self-evolving Operating System Layer
2026-09-18
Making AI-Assisted Claims Independently Challengeable: Publication Authority and a Protocol for Falsifiable Publication Records
2026-09-17
EvolveTrade: Experience-Driven Policy Refinement for Self-Evolving LLM Trading Agents
2026-09-17
One Color Preprocessing Improves DSATUR
2026-09-17
Physics-Constrained Digital Twins for Sensor Integrity in Urban Pedestrian Flow: Detecting Stealthy False Data Injection with Conformal Guarantees
2026-09-17
Optimal Pruning for Neural Architectures using Fisher Information Distances
2026-09-16
Safe Error Correction for Language Models: Frozen-Base Adjustment with Capability Preservation
2026-09-16
GPEvac: GNN-Based PPO for Adaptive Evacuation Routing During Shooting Events
2026-09-16
Position: AI Is Not Ready for Strategic Conflicts
2026-09-16
ZGCM-1: A Fully Open and Extremely Efficient Foundation Model for Math and Agentic Search
2026-09-15
Converge Then Diversify: Decoupling Convergence and Diversity in Multi-Objective Bayesian Optimisation
2026-09-15
Generalized Agent Iteration: One Formal Framework for Iterative Policy Improvement and Recursive Self-Improvement
2026-09-15
Vibe Patenting: Evaluating LLM Judges for Professional Patent-Drafting Agents
2026-09-15
Probabilistic Focal Search: Accelerating Bounded-Suboptimal Search via Lower-Bound Advancement
2026-09-12
Automating Quadratic Unconstrained Binary Optimization (QUBO) Formulation Generation from Natural Language
2026-09-12
A Multi-Stage Rule-Chaining Framework for Compositional and Interpretable Cognitive Reasoning
2026-09-12
Understanding LoRA Rank Trade-offs in Diffusion Model Fine-Tuning
2026-09-12
OpenDiscoveryTrace: Process Traces for Evaluating AI Scientist Workflows
2026-09-11
Adaptive Entangled Game Modules in Artificial General Intelligence
2026-09-11
Subagents vs Agent Skills: Executing Reusable Knowledge for Long-Horizon Agentic Tasks
2026-09-11
Gradland: On Phenomenal Experience, Differentiated Across Many Dimensions
2026-09-11
Beyond Right and Wrong: Evaluating Second-order Social Reasoning in Large Language Models
2026-09-10
CriticGen: Generation-Aware Evaluation as Actionable Feedback
2026-09-10
When Does Memory Help? A Cost-Aware Evaluation of Long-Term Memory in Tool-Using LLM Agents
2026-09-10
AutoFyn Technical Report: Non-Parametric Expert Iteration for Long-Horizon Agents
2026-09-10
EXAONE Forecast for Finance
2026-09-07
From Matching Models to Recruiting Agents: A Systematized Narrative Review of AI Recruitment Systems, Evaluation, and Governance
2026-09-07
Harbor Adapters and Harbor-Index: Infrastructure and a Curated Meta-Dataset for Large-Scale Agentic Evaluation
2026-09-07
Data-Optimized Contingency Screening: A Machine Learning Approach to Power System Security
2026-09-07
Structure and Implementation of New Practical English Textbooks Driven by Artificial Intelligence
2026-09-04
MasterControl Seventeen Every Time
2026-09-04
Speculative Macro Commit for Faster Tool-Using Agents
2026-09-04
Fresh Memory, Stale Plans: Dependency-Scoped Validation for Distributed LLM-Agent Memory
2026-09-04
EvalDetectBench: A Benchmark for Measuring Evaluation Awareness in Frontier Language Models
2026-09-03
Meta-ethics and AI: exploring the novel meta-ethical questions in the era of AI
2026-09-03
When Can a Machine Trust a Statute? A Survival Certificate for Machine-Extracted Legal Logic
2026-09-03
When Does Information Sharing Improve Decentralized Discovery? Aggregation, Independent Rescue, and Equilibrium Selection
2026-09-03
HyperWorld: Hypergraph-Structured State Serialization Improves Learned Textual World Models
2026-09-02
I-CARE: Analysis of interference-related phenomena in a controllable, diverse and representative unlearning setting for text-to-image models
2026-09-02
Discrete-Time MDP Modeling for Multi-Item Capacitated Lot Sizing with Stochastic Demand Timing
2026-09-02
Incremental Risk Assessment of Progressive Elder Financial Scams via Instruction-Tuned Small Language Models
2026-09-02
DS-Lighting: Making Agent Harnesses Explicit for Data-Science Automation
2026-09-01
Expert-validated STEM QA
2026-09-01
A collective capability boundary in frontier large language models on guideline-conformant and case-specific oncology decision-making
2026-09-01
Statutory AI: Aligning Large Language Models With Legal Norms
2026-09-01
EduRiskX: A Neuro-Symbolic Framework with F-Logic Reasoning for Early Academic Risk Prediction
2026-08-28
Standalone LLM and a Pre-specified Agentic Pipeline for Explaining ICU Mortality Predictions: a Feasibility Study on the eICU Demo Dataset
2026-08-28
Large Models for Battery Prognostics and Health Management: A Review and Future Roadmap
2026-08-28
PICasso: An AI-Enabled Design Framework for Autonomous Optimization of Silicon Photonic Devices
2026-08-28
RENDER: Controlling Reader-Facing Evidence in LLM Memory Evaluation
2026-08-27
ESQ-Bench: A Multi-Tier Enterprise Oracle Benchmark for Evaluating NL2SQL Dialect Generalization and Silent Semantic Divergence
2026-08-27
LLM Agents Perform Controlled Experiments Using Simulation Models
2026-08-27
A survey detection channel overrides the pixels in an astronomical foundation model, and biases tomographic mean redshifts
2026-08-27
Ranked #21 of 463
by items recorded in our collection · Nearby:
↑ cs.CL
·
↓ OpenAI
Get 4 AI signals a day by email — free.
Subscribe free →
See all plans →
Get 4 AI signals a day by email — free
Subscribe free
All models
·
All repos
·
By company
·
Daily editions