AiAnyTool - Best AI Tools Directory and Artificial Intelligence Software Hub Logo
Loading theme toggle

Real-Time Coverage

AI News Today

Live

9571 stories from 30+ sources, refreshed continuously.

arXiv cs.LGResearch

Decentralized Best-Response-Based Learning in Two-Player Zero-Sum Stochastic Games: A Finite-Sample Analysis

arXiv cs.LGResearch

Normalizing Flows are Capable Models for Continuous Control

arXiv cs.LGResearch

Reinforcement Fine-Tuning of Flow-Matching Policies for Vision-Language-Action Models

arXiv cs.LGResearch

Spurious Rewards Paradox: Mechanistically Understanding How RLVR Activates Memorization Shortcuts in LLMs

arXiv cs.LGResearch

Learning State-Tracking from Code Using Linear RNNs

arXiv cs.LGResearch

Use What You Know: Causal Foundation Models with Partial Graphs

arXiv cs.LGResearch

BrepCoder: A Unified Multimodal Large Language Model for Multi-task B-rep Reasoning

arXiv cs.LGResearch

Training-Free Generation of Protein Sequences from Small Family Alignments via Stochastic Attention

arXiv cs.LGResearch

NASimJax: A GPU-Accelerated Policy Learning Framework for Penetration Testing

arXiv cs.LGResearch

Learning from Equivalence Queries, Revisited

arXiv cs.LGResearch

Chebyshev Policies and the Mountain Car Problem: Reinforcement Learning for Low-Dimensional Control Tasks

arXiv cs.CL (NLP)Research

Orthogonal Hierarchical Decomposition for Structure-Aware Table Understanding with Large Language Models

arXiv cs.CL (NLP)Research

When Actions Go Off-Task: Detecting and Correcting Misaligned Actions in Computer-Use Agents

arXiv cs.CL (NLP)Research

ReportLogic: Evaluating Logical Quality in Deep Research Reports

arXiv cs.CL (NLP)Research

Embarrassingly Simple Self-Distillation Improves Code Generation

arXiv cs.CL (NLP)Research

From Guessing to Placeholding: A Cost-Theoretic Framework for Uncertainty-Aware Code Completion

arXiv cs.CL (NLP)Research

ScheMatiQ: From Research Question to Structured Data through Interactive Schema Discovery

arXiv cs.CL (NLP)Research

Peer-Preservation in Frontier Models

arXiv cs.CL (NLP)Research

Why Are Some Emotions Harder for LLMs? Uncovering the Causal Mechanisms of Emotion Inference via Sparse Autoencoders

arXiv cs.CL (NLP)Research

Not All Proofs Are Equal: Evaluating LLM Proof Quality Beyond Correctness

arXiv cs.CL (NLP)Research

Weak-to-Strong Elicitation via Mismatched Wrong Drafts

arXiv cs.CL (NLP)Research

The Annotation Scarcity Paradox in Low-Resource NLP Evaluation: A Decade of Acceleration and Emerging Constraints

arXiv cs.CL (NLP)Research

See, Infer, Intervene: Proactive World Modeling for Goal-Oriented Social Intelligence

arXiv cs.CL (NLP)Research

Does AI Reviewer See the Full Picture? Attacking and Defending Multimodal Peer Review

arXiv cs.CL (NLP)Research

Learning from the Self-future: On-policy Self-distillation for dLLMs

arXiv cs.AIResearch

LiMoDE: Rethinking Lifelong Robot Manipulation from a Mixture-of-Dynamic-Experts Perspective

arXiv cs.AIResearch

Statistical and Structural Approaches to Algorithmic Fairness

arXiv cs.AIResearch

CyberChainBench: Can AI Agents Secure Smart Contracts Against Real-World On-Chain Vulnerabilities?

arXiv cs.AIResearch

Lacuna: A Research Map for Machine Learning

arXiv cs.AIResearch

A multi-task spatiotemporal deep neural network for predicting penetration depth and morphology in laser welding

arXiv cs.AIResearch

TEMPO-Diffusion: Temporally Exposed Malicious Poisoning of Diffusion Models

arXiv cs.AIResearch

SSM Adapters via Hankel Reduced-order Modeling: Injection Site Determines Task Suitability in Long-Context Fine-Tuning

arXiv cs.AIResearch

The Red Queen G\"odel Machine: Co-Evolving Agents and Their Evaluators

arXiv cs.AIResearch

Parametric Generalized Adaptive Moment Features (PG-AMF) for Bearing Fault Diagnosis and Machine Health Monitoring

arXiv cs.AIResearch

EVOM: Agentic Meta-Evolution of Actor-Critic Architectures for Reinforcement Learning

arXiv cs.AIResearch

SOLAR: AI-Powered Speed-of-Light Performance Analysis

arXiv cs.AIResearch

Sampling sea state using a diffusion model

arXiv cs.AIResearch

Beyond Feedforward Networks: Reentry Neural Systems as the Fundamental Basis of Subjecthood and Intrinsic Safety of Next-Generation AGI

arXiv cs.AIResearch

CoStream: Composing Simple Behaviors for Generalizable Complex Manipulation

arXiv cs.AIResearch

Play2Perfect: What Matters in Dexterous Play Pretraining for Precise Assembly?