AiAnyTool - Best AI Tools Directory and Artificial Intelligence Software Hub Logo
Loading theme toggle

Real-Time Coverage

AI News Today

Live

24764 stories from 30+ sources, refreshed continuously.

arXiv cs.LGResearch

Data-Efficient Adaptation of LLMs via Attention Head Reweighting

arXiv cs.LGResearch

PUe: Biased Positive-Unlabeled Learning Enhancement by Causal Inference

arXiv cs.CL (NLP)Research

FixItFlow: Automated Troubleshooting Guide Generation from Cloud Incidents

arXiv cs.CL (NLP)Research

Ask Before You Diagnose: Safe-Psych, a Sequential Evaluation Benchmark for LLMs in Psychiatry

arXiv cs.CL (NLP)Research

The Perplexity Trap: When Patent Law Makes Human Writing Look Like AI

arXiv cs.CL (NLP)Research

Do LLMs Need Architectural Changes for Simultaneous Speech Translation? A Prefix-to-Prefix Data Driven Approach

arXiv cs.CL (NLP)Research

What Models Express, Suppress, and Resist: Auditing Open-Weight LLMs with Persona Vectors

arXiv cs.CL (NLP)Research

Text2Sign: A Single-GPU Diffusion Baseline for Text-to-Sign Language Video Generation

arXiv cs.CL (NLP)Research

RAGthoven at SemEval-2026 Task 1: A Multi-Stage Pipeline Walks Into a Benchmark and Barely Clears the Bar

arXiv cs.CL (NLP)Research

Adaptive Filtering of the KV Cache: Diagnosing and Correcting Structural-Role Bias in LLM Inference

arXiv cs.CL (NLP)Research

GSM-Plus-BN: A Perturbation-Based Benchmark for Bangla Mathematical Reasoning in Large Language Models

arXiv cs.CL (NLP)Research

Discourse-Aware Policy Analysis with Argumentation: A Hybrid LLM-Symbolic Framework for Disaster Governance

arXiv cs.CL (NLP)Research

Meta-Learning Preferences for Multilingual LLM Alignment

arXiv cs.CL (NLP)Research

Evaluation Ability Does Not Imply Optimization Utility: LLM-as-a-Judge Signals in Closed-Loop Table Recognition

arXiv cs.CL (NLP)Research

GFlowRL: Scaling Distribution-Matching RL to Large Language Models

arXiv cs.CL (NLP)Research

Demystifying On-Policy Distillation: Roles, Pathologies, and Regulations

arXiv cs.CL (NLP)Research

Exploring Post-Training Alignment of Small Language Models for Biomedical Data-to-Text Generation: A Case Study of Medication Leaflet

arXiv cs.CL (NLP)Research

When Rubrics Change: Cross-Rubric Generalization for Critical Thinking Essay Scoring

arXiv cs.CL (NLP)Research

DevicesWorld: Benchmarking Cross-Device Agents in Heterogeneous Environments

arXiv cs.CL (NLP)Research

MyAG: A Graph-Based Framework for Designing and Analyzing Composable LLM Agent Systems

arXiv cs.CL (NLP)Research

Cost-Pragmatic Quality Gating and Selection-Fusion Multi-Model Combiners for BioASQ Phases A+ and B

arXiv cs.CL (NLP)Research

Graded Entity-Familiarity Readouts in Language Models: Polish Adaptation, Cross-Language Robustness, and Refusal Steering

arXiv cs.AIResearch

UESF-Bench: Benchmarking and Probing for Unified Embodied Seeking and Following

arXiv cs.AIResearch

Explaining Reinforcement Learning Agents via Inductive Logic Programming

arXiv cs.AIResearch

When Bots Join the Team: Bot Adoption and the Institutional Fabric of Open-Source Software Projects

arXiv cs.AIResearch

AgentCompass: A Unified Evaluation Infrastructure for Agent Capabilities

arXiv cs.AIResearch

CAVA: Canonical Action Verification and Attestation for Runtime Governance of Agentic AI Systems

arXiv cs.AIResearch

Experience Memory Graph: One-Shot Error Correction for Agents

arXiv cs.AIResearch

AIMO Interpretability Challenge

arXiv cs.AIResearch

A Self-Evolving Agent for Longitudinal Personal Health Management

arXiv cs.AIResearch

Do Agent Optimizers Compound? A Continual-Learning Evaluation on Terminal-Bench 2.0

arXiv cs.AIResearch

AI-accelerated End-to-End Framework for Rapid Professional Upskilling

arXiv cs.AIResearch

Earthquaker-AI: A Retrieval-Augmented Generation Framework with Rubric-Based Assessment for Primary School Earthquake Education

arXiv cs.AIResearch

Deep Interaction: An Efficient Human-AI Interaction Method for Large Reasoning Models

arXiv cs.AIResearch

Designing Safety-Constrained LLM Systems for Public Health Information Access

arXiv cs.AIResearch

Final Authority in AI Governance: Frontier-Provider Sovereignty and Action-Centered Deployer Governance

arXiv cs.AIResearch

LessonBench-V1: A Benchmark Dataset for Evaluating AI Lesson Generation Agents

arXiv cs.AIResearch

Beyond Backbone Backpropagation: A Decoupled Strategy for Efficient Transfer Learning

arXiv cs.AIResearch

Autonomous UAV Route Planning for Coverage Maximization in Environmental Monitoring: A Systematic Literature Review

arXiv cs.AIResearch

Compaction as Epistemic Failure: How Agentic LLM Tools Fabricate Confirmed Results from Killed Processes