AI News

⚡ 5 minutes ago
1
1
Failures Reveal What Metrics Miss: An Evidence-Driven Agent for Recursive Refinement of ECG Classifiers (arxiv.org)
2
1
PANOPTICON: A PII-Based Assemblage of Naturalistic Output Tokens for Investigating Privacy Leakage Within LLM Context Window (arxiv.org)
3
1
The Scaffold Effect in Coding Agents: Harness Choice as a Hidden Variable in Coding-Agent Evaluation (arxiv.org)
4
1
TriSP: Tri-Signal Structured Pruning for Large Language Models (arxiv.org)
5
1
AI-Assisted Causal Inference and Mediation Analyses of Environmental and Psychosocial Determinants of Subjective Cognitive Difficulties in the All of Us Research Program (arxiv.org)
6
1
Plato-Bio: verification-first biological novelty screening with temporal rediscovery and structural benchmarks (arxiv.org)
7
1
Traceable LLM Reasoning for Fake-Order Fraud Detection (arxiv.org)
8
1
Loss-Aware Feature-Map Pruning in Convolutional Neural Networks Using Multi-Armed Bandits (arxiv.org)
9
1
Falsifiable Commitment Planning for Self-Correcting Web Agents (arxiv.org)
10
1
Retrieval-Augmented Generation of Ontologies from Relational Databases (arxiv.org)
11
1
CSV-Decode: Certifiable Sub-Vocabulary Decoding for Efficient Large Language Model Inference (arxiv.org)
12
1
TLRNet: Estimating Individual Treatment Effect based on Local Information and Single Learner Structure (arxiv.org)
13
1
Debiased Machine Learning: Identification, Estimation, and Shape Constraints (arxiv.org)
14
1
MIITA: Memory-Induced Inference-Time Adaptation for Continual Learning with Small Language Models (arxiv.org)
15
1
EmotionAI: A Privacy-Preserving Computational Intelligence Pipeline for Speech-Emotion-Grounded Conversational Analysis (arxiv.org)
16
1
Multi-Objective Structured Pruning of LLMs for Latency and Model Size Optimization (arxiv.org)
17
1
Distributional Split Criteria for Random Forests: Extensions, Shrinkage, and the Robustness of Mean Splitting (arxiv.org)
18
1
Lions and Muons: Optimization via Stochastic Frank-Wolfe under Heavy-Tailed Noise (arxiv.org)
19
1
Robust Conformalized Selection with Noisy Responses (arxiv.org)
20
1
Modeling Memory-Dependent Reliability of LLMs: A Hidden Markov Model (arxiv.org)
21
1
Amortized Bayesian Causal Discovery of Extended Factor Graphs (arxiv.org)
22
1
Two-Timescale Hierarchical Reinforcement Learning for Resilient Operations (arxiv.org)
23
1
Agent-UCT: Upper Confidence Bounds Applied to Trees for Agentic Workflow Optimization with Cost-Awareness (arxiv.org)
24
1
Covariance-Boosted Gaussian Processes for Spatiotemporal Irregularities (arxiv.org)
25
1
GNN-based Multi-Agent Control of Traffic Shockwaves in Sparse Vehicular Ad-hoc Networks (arxiv.org)
26
1
Lexical discovery in unknown environments orchestrated by Large Language Models (arxiv.org)
27
1
Scoping Review of AI, Metrology, and ESG in the Semiconductor Sector: Implications for Safe and Sustainable by Design (SSbD) (arxiv.org)
28
1
MM-ShiftKV: Decode-Aware Prefill-Stage KV Selection for Multimodal Large Language Models (arxiv.org)
29
1
VLASH: Real-Time VLAs via Future-State-Aware Asynchronous Inference (arxiv.org)
30
1
MIME: Multimodal Interactive Motion Encoder (arxiv.org)
31
1
TRUAV: Distributed Multi-Agent Reinforcement Learning for Trajectory Planning and Routing Enhancement in UAV-Aided IoT-Enabled VANETs (arxiv.org)
32
1
What CLIP Knows but Cannot Say: Recovering Negation from Frozen Intermediate Features (arxiv.org)
33
1
Understanding Human-like Solutions in Combinatorial Optimization via Learning and Search (arxiv.org)
34
1
TriShieldRAG: A Three-Ring Defense-in-Depth Framework Against Knowledge Corruption in Retrieval-Augmented Generation (arxiv.org)
35
1
RoleMix: Unifying Sequential and Non-Sequential Features via Semantic Tokenization for Post-Click Conversion Rate Prediction (arxiv.org)
36
1
HiLLTS: Zero-Shot Hierarchical LLM-Guided Traffic Signal Control for Sustainable Transportation (arxiv.org)
37
1
Decision trees, Frobenius traces, and Weierstrass coefficients of elliptic curves (arxiv.org)
38
1
Source-Aware Reranking for Retrieval-Augmented Generation: A Reliability Prior Approach (arxiv.org)
39
1
ARdena: Scenario-driven control of real-time LLM agents (arxiv.org)
40
1
FedSLIM: Privacy-Preserving Federated MDL-Based Descriptive Pattern Mining Across Data Silos (arxiv.org)
41
1
A Fixed-Effects Causal Forest for Staggered Adoption, with an Application to Medicaid Expansion (arxiv.org)
42
1
Epistemic Norms for AI Safety and Alignment Research (arxiv.org)
43
1
ADAGE: A Language-Agnostic Pipeline for Analogical Reasoning Evaluation (arxiv.org)
44
1
GeoDecider: An Evidence-Grounded Agent for Geological Interpretation via Deliberative Reasoning (arxiv.org)
45
1
From RLVR to RLSVR: Task Transformation Induces Self-Verifiable Rewards for Open-Ended LLM Self-Improvement (arxiv.org)
46
1
CodeEvo: Interaction-Driven Synthesis of Code-centric Data through Hybrid and Iterative Feedback (arxiv.org)
47
1
Imprompt: A Language Framework for Prompt Programming (arxiv.org)
48
1
Realizing Scaling Laws in Recommender Systems: A Foundation-Expert Paradigm for Hyperscale Model Deployment (arxiv.org)
49
1
HydroAgent: Formalizing Forecaster Expertise into Skill-Orchestrated Flood Forecasting Workflows (arxiv.org)
50
1
Nanbeige4.2-3B: Unlocking Agentic Capabilities in a Compact Model (arxiv.org)