AiAnyTool - Best AI Tools Directory and Artificial Intelligence Software Hub Logo
Loading theme toggle

Real-Time Coverage

AI News Today

Live

24856 stories from 30+ sources, refreshed continuously.

arXiv cs.CL (NLP)Research

Robust Summarization of Doctor-Patient Conversations: TalTech Systems for the Beyond Transcription Challenge

arXiv cs.CL (NLP)Research

EvolvingWorld: An Open-Schema Framework for Co-Evolving Role-Play Agents and World Model in Interactive Literary World

arXiv cs.CL (NLP)Research

Debate-on-Graph: Reliable and Adaptive Reasoning of Large Language Model on Uncertain Knowledge Graph

arXiv cs.CL (NLP)Research

Safety That Does Not Transfer: Cross-Lingual Clinical Correctness Drift in Deployable Medical Language Models

arXiv cs.CL (NLP)Research

Team DACTYL at PAN 2026: Bayesian Data Mixing and Empirical X-risk Minimization for AI-text Detection

arXiv cs.CL (NLP)Research

The Librarian Who Refused to Code: Model-Dependent Identity Enactment in LLM Code Generation

arXiv cs.CL (NLP)Research

Token-Level Off-Policy Learning for Faithful Generation Under Distribution Shift

arXiv cs.CL (NLP)Research

Oracle Gap and Signal Fidelity: A Fixed-Pool Diagnostic for Test-Time Collaboration

arXiv cs.CL (NLP)Research

C$^2$KV: Compressed and Composable KV Cache Reuse for Efficient LLM Inference

arXiv cs.CL (NLP)Research

Large Language Models for Citation Function Classification

arXiv cs.CL (NLP)Research

When a Name Is Not a Name: A Benchmark Dataset and Distilled Reasoning for Culturally Entangled Bangla Homographs in Low-Resource LLMs

arXiv cs.CL (NLP)Research

Zero Hallucination, by Construction: Hallucination-Aware Layered Oversight for Trustworthy Enterprise AI

arXiv cs.CL (NLP)Research

DeLIVeR: Decomposed Learning for Information-grounded Veracity Recognition via Reinforced Knowledge Graph Exploration

arXiv cs.CL (NLP)Research

What Transfers Under Source Shift? Definitions, Examples, and Fine-Tuning for Climate Disclosure Classification

arXiv cs.CL (NLP)Research

An Early Warning of Emerging Biosecurity Risks in Frontier LLMs

arXiv cs.CL (NLP)Research

Pancasila-Dilemmas: Evaluating Large Language Models on Indonesian Human Value Dilemmas Grounded in Pancasila

arXiv cs.CL (NLP)Research

Modeling turn-taking with distant viewing: investigating silence thresholds in human and AI-generated discourse

arXiv cs.CL (NLP)Research

VDAR-Router: Adaptive LLMs Routing via Verbalized Query Difficulty Analysis Retrieval

arXiv cs.CL (NLP)Research

How Does Alignment Tuning Shape Representations of Sycophancy and Related Cue-Induced Biases in LLMs?

arXiv cs.CL (NLP)Research

VEHBench: A Stage-Local Diagnostic Benchmark for LLM-Assisted Vibration Energy Harvester Design

arXiv cs.AIResearch

When LLMs Over-Answer: Measuring and Mitigating Quality Issues in LLM-Based Hardware Description Language Question Answering

arXiv cs.AIResearch

Bridging the Information Gap: Semantic Densification and Hindsight Distillation for Cold-Start Prediction

arXiv cs.AIResearch

Otap:Structure-Aware Optimal Transport for Evaluating Planning and Execution in Agent Trajectories

arXiv cs.AIResearch

A Diagnostic Framework for AI Agent Behavior

arXiv cs.AIResearch

A Systematic Evaluation of Trajectory Data Curation for LoRA Fine-Tuning of Code Agents

arXiv cs.AIResearch

Constrained Path Reasoning: Measuring When Committed Stages Earn Their Cost

arXiv cs.AIResearch

LenGuard-GPC: Length Guarding with Guided-Prompt Consistency for Spatial Reasoning Reinforce Learning

arXiv cs.AIResearch

An Explicit World Model Based on Data-First Ontology: DaoQL Multimodal Storage Validation and Counterfactual Reasoning Evaluation

arXiv cs.AIResearch

Lossless but Not Free: An Empirical Anatomy of Speculative Decoding on Consumer Hardware

arXiv cs.AIResearch

Agentic ERP: Multi-Agent Large Language Model Architecture for Autonomous Enterprise Resource Planning

arXiv cs.AIResearch

DeeperRadar: End-to-End MIMO Radar Design and Multi-Modal Fusion for Autonomous Vehicle Perception

arXiv cs.AIResearch

Self-Modifying Lean Proof Agents with Verifier-Grounded Benchmark Coevolution

arXiv cs.AIResearch

Quantifying Diversity of Thought: A Predictive Law of Weighted LLM Ensemble Lift

arXiv cs.AIResearch

Intermittent Control Is Not Diluted Control: A Switching Effect in Artificial Agency

arXiv cs.AIResearch

Empirical Grounding Improves the Realism of LLM Agents Simulating Human Behavior During Disruptions

arXiv cs.AIResearch

Can AI Agents Really Complete RTL-to-GDS? Lessons from Benchmarking Tool-Interactive EDA Workflows

arXiv cs.AIResearch

Retain or Consolidate? Budget-Dependent Operator Selection for Language Agent Memory

arXiv cs.AIResearch

Why Does Feedback-Augmented Self-Distillation Fail to Improve Retrieval-Interleaved Search Agents?

arXiv cs.AIResearch

Reinforcement Learning: From Algorithms To Foundation Models

arXiv cs.AIResearch

ZifaMem: Structured Memory for Persona, Preference, and Emotional Continuity in AI Companions