AI News

⚡ 8 minutes ago
1
1
SimBEV2X: A Large-Scale Dataset and Data Generation Tool for Multi-Task Vehicle-to-Everything Cooperative Perception (arxiv.org)
2
1
Synthetic Scenario Generation for Evaluation of Industry 4.0 Agents (arxiv.org)
3
1
MedLoCoMo: A Long-Context Multi-Session Medical Dialogue Benchmark for Large Language Models (arxiv.org)
4
1
Lexical discovery in unknown environments orchestrated by Large Language Models (arxiv.org)
5
1
Multi-Agent Privacy Game in Federated Learning: A Unified Mean-Field View (arxiv.org)
6
1
Failures Reveal What Metrics Miss: An Evidence-Driven Agent for Recursive Refinement of ECG Classifiers (arxiv.org)
7
1
BoneAgeTW2: Automated Skeletal Maturation Assessment via the Tanner-Whitehouse 2 Method, Deep Learning, and Clinical Report Generation with Distribution Curves (arxiv.org)
8
1
DRP-FLR: Data-Driven Assessment of Demand Response Potential for Flexible Load Regulation in Smart Grids (arxiv.org)
9
1
Explainable Reinforcement Learning via Physics-Aware Policy Distillation (arxiv.org)
10
1
Label-free Industrial Fault Detection via Adversarial Inverse Reinforcement Learning: A System for Run-to-Failure Prognostics (arxiv.org)
11
1
DSTFView: Multi-View Cloud-Edge Workload Forecasting with Dual-Input Spatio-Temporal-Frequency Modeling (arxiv.org)
12
1
Training Language Models to Cooperate with Inference-Time Controllers (arxiv.org)
13
1
Epistemic Norms for AI Safety and Alignment Research (arxiv.org)
14
1
Too much evidence, too little time: From text to actionable recommendations through multi-objective evidence reasoning (arxiv.org)
15
1
MixQuant: Adaptive Mixed-Precision Quantization for Large Language Models (arxiv.org)
16
1
When Can You Correct Distribution Drift in Temporal Graph Generation? A Sharpening--Drift Tension and an Impossibility for Observation-Based Correction (arxiv.org)
17
1
ARdena: Scenario-driven control of real-time LLM agents (arxiv.org)
18
1
MM-ShiftKV: Decode-Aware Prefill-Stage KV Selection for Multimodal Large Language Models (arxiv.org)
19
1
AI-Assisted Causal Inference and Mediation Analyses of Environmental and Psychosocial Determinants of Subjective Cognitive Difficulties in the All of Us Research Program (arxiv.org)
20
1
A Vocabulary for Multi-Agent Automated Research Systems (arxiv.org)
21
1
MioFFAn: an Annotation Software for Formula Formalization with LLM Automation Capabilities (arxiv.org)
22
1
Falsifiable Commitment Planning for Self-Correcting Web Agents (arxiv.org)
23
1
TriShieldRAG: A Three-Ring Defense-in-Depth Framework Against Knowledge Corruption in Retrieval-Augmented Generation (arxiv.org)
24
1
MPR-CiteG: Enhancing RAG with Multi-Portfolio Retrieval and Citation-Grounded Generation (arxiv.org)
25
1
A Frozen 12B Beats Frontier Models on Verified Work: 100% Accuracy, 0 Tokens, Bit-Exact, Forever (arxiv.org)
26
1
QFoldAgent: An Autonomous Quantum Optimization Multi-Agent System for Protein Structure Prediction (arxiv.org)
27
1
Delegation Intelligence in Deep Search: A Controllable Framework for Disentangled Capability Diagnosis (arxiv.org)
28
1
Making Mathematical Knowledge Explainable, Accessible and Interoperable Through Large Language Model Integration (arxiv.org)
29
1
Task-Conditional Faithfulness Auditing of Multimodal LLMs for Grid Diagnosis (arxiv.org)
30
1
Numerical Investigation of Sequence Modeling Theory using Controllable Memory Functions (arxiv.org)
31
1
Loong: Synthesize Long Chain-of-Thoughts at Scale through Verifiers (arxiv.org)
32
1
From Execution to Capability: Scientific Experience Consolidation via Procedural Knowledge Synthesis (arxiv.org)
33
1
Integrating Factual and Normative Industrial Knowledge via Constraint-Aware Graph Attention for Process Plan Recommendation (arxiv.org)
34
1
TRACE-CTI: Auditable Post-Extraction Governance of TTP Claims with Knowledge Graphs (arxiv.org)
35
1
A New Kind of Adversarial Example: Measuring the Human-Model Gap, and Its Relationship to OOD Detection (arxiv.org)
36
1
Discrepancy-Rounded Fair Bandits with Static and Time-Varying Exposure Floors (arxiv.org)
37
1
ParBench: A Benchmark for Reliable Evaluation of LLM Parallel Code Translation (arxiv.org)
38
1
VecTree-RAG: An Agentic Retrieval-Augmented Generation Framework Combining Vector and Tree Retrieval for Efficiency and Accuracy (arxiv.org)
39
1
Online Policy Evaluation for MDPs with Dynamic UBSR Measures (arxiv.org)
40
1
Sampling Decisions: Exact Path-Space Correction, Prior Cancellation and Local-Boltzmann Guidance (arxiv.org)
41
1
Distribution-Specific Curvature Control with Finite-Sample Guarantees for Open-Weight Safety (arxiv.org)
42
1
Spectral-Aware Analytic Class-Incremental Learning for Long-Tailed Distributions (arxiv.org)
43
1
PYPM-GGD: Pitman-Yor Process Mixture with Generalized Gaussian Density using ADAM (arxiv.org)
44
1
CausAdv: A Causal-based Framework for Detecting Adversarial Examples (arxiv.org)
45
1
Verbalized Particle Posterior: Bayesian Inference over Natural Language Hypotheses (arxiv.org)
46
1
StageGuard: Physiologically Constrained Sleep Staging (arxiv.org)
47
1
All in One: Generative Modeling as Mean-Field Game Design (arxiv.org)
48
1
ObsDriveBench: Benchmarking Multimodal Understanding under Adverse Weather with Observability Awareness (arxiv.org)
49
1
LLM-SoccerArena: Benchmarking LLMs on Real-World Predictions in Sports (arxiv.org)
50
1
Verification-Notebook Learning for Source-Aware Multimodal Misinformation Detection (arxiv.org)