AI News

⚡ 12 minutes ago
1
1
Temporal Context Reinstatement Drives Episodic-Like Order Memory in Long-Context Language Models (arxiv.org)
2
1
PhononBench-MP40: a spectrum-resolved benchmark dataset for phonon stability (arxiv.org)
3
1
Multi-Objective Structured Pruning of LLMs for Latency and Model Size Optimization (arxiv.org)
4
1
Kinship Verification through a Forest Neural Network (arxiv.org)
5
1
The Scaffold Effect in Coding Agents: Harness Choice as a Hidden Variable in Coding-Agent Evaluation (arxiv.org)
6
1
TriSP: Tri-Signal Structured Pruning for Large Language Models (arxiv.org)
7
1
SimBEV2X: A Large-Scale Dataset and Data Generation Tool for Multi-Task Vehicle-to-Everything Cooperative Perception (arxiv.org)
8
1
Long-Tailed Medical Image Classification (arxiv.org)
9
1
BERT-based Models vs. Large Language Models for Low-Resource Named Entity Recognition: A Comparative Study on Marathi (arxiv.org)
10
1
A Frozen 12B Beats Frontier Models on Verified Work: 100% Accuracy, 0 Tokens, Bit-Exact, Forever (arxiv.org)
11
1
Attribution and Uncertainty Behavior of Learned Residual Gyro Correction for Gyro-Stellar Estimation (arxiv.org)
12
1
DP-IVON-Gradsq: Differentially Private Squared-Gradient Improved Variational Online Newton (arxiv.org)
13
1
Agent-UCT: Upper Confidence Bounds Applied to Trees for Agentic Workflow Optimization with Cost-Awareness (arxiv.org)
14
1
TRE: Training-Free Hallucination Detection for Diffusion Language Models (arxiv.org)
15
1
MIITA: Memory-Induced Inference-Time Adaptation for Continual Learning with Small Language Models (arxiv.org)
16
1
DeepLens Diagnosis Agent: Agentic Workflow Design Lets a Small Reasoning Model Compete with Frontier LLMs (arxiv.org)
17
1
Causal-TS: A Python Library for Causal Discovery in High-Dimensional and Nonstationary Time Series (arxiv.org)
18
1
Loss-Aware Feature-Map Pruning in Convolutional Neural Networks Using Multi-Armed Bandits (arxiv.org)
19
1
CodexGraph: Bridging Large Language Models and Code Repositories via Code Graph Databases (arxiv.org)
20
1
MolCryst-MLIPs: A Machine-Learned Interatomic Potentials Database for Molecular Crystals (arxiv.org)
21
1
AI-Assisted Causal Inference and Mediation Analyses of Environmental and Psychosocial Determinants of Subjective Cognitive Difficulties in the All of Us Research Program (arxiv.org)
22
1
CALMRec: Causally Aligned Language Memory for Long-Horizon Recommendation (arxiv.org)
23
1
An Unofficial FastLAS Tutorial: A Programmer's Guide (arxiv.org)
24
1
TRACE-CTI: Auditable Post-Extraction Governance of TTP Claims with Knowledge Graphs (arxiv.org)
25
1
Generative Artificial Intelligence (GenAI) to convert images of queuing networks into verifiable simulation models: an open-weight LLM workflow approach (arxiv.org)
26
1
Epistemic Norms for AI Safety and Alignment Research (arxiv.org)
27
1
To Erase, or Not to Erase: Robust Training-Free Concept Erasure with Preservation aware Adaptive Ranked Subspace Expansion (arxiv.org)
28
1
Action from Adjacent Set in Physical Space Outperforms the Best Prediction in World Models (arxiv.org)
29
1
Plans Work in Mysterious Ways: Evaluating a Plan Mode for Spreadsheet Agents (arxiv.org)
30
1
Failures Reveal What Metrics Miss: An Evidence-Driven Agent for Recursive Refinement of ECG Classifiers (arxiv.org)
31
1
ATLAS: Automated Approximation of Transformers for Efficient Homomorphic Inference in One Hour (arxiv.org)
32
1
Variational-Ising-Attention (VIA):TailoredAttentionMattersfor Science (arxiv.org)
33
1
Stacking the Deck: Tunable Trainability in Stacked LCUs (arxiv.org)
34
1
Simulating Tenant Responses to Energy Policy Interventions with Transaction-Cost-Aware LLM Age (arxiv.org)
35
1
Approximate reservoir computing with a semiconductor laser for reducing energy consumption (arxiv.org)
36
1
LLM-SoccerArena: Benchmarking LLMs on Real-World Predictions in Sports (arxiv.org)
37
1
MIME: Multimodal Interactive Motion Encoder (arxiv.org)
38
1
Training Language Models to Cooperate with Inference-Time Controllers (arxiv.org)
39
1
DSCH-Loss: A Dynamic Semantic Channel Objective for Deep Semantic Hashing (arxiv.org)
40
1
Visual Token Compression Enhances Robustness of MLLMs (arxiv.org)
41
1
A New Kind of Adversarial Example: Measuring the Human-Model Gap, and Its Relationship to OOD Detection (arxiv.org)
42
1
Structure over Depth: A Single-Block Spatio-Temporal Transformer for Multi-Entity Reasoning (arxiv.org)
43
1
Obliviate: Efficient Unlearning in Recommender Systems (arxiv.org)
44
1
Frustratingly Simple Black-Box Adaptation of Language Models via Logit Bias (arxiv.org)
45
1
AIR-BENCH Live: An Evolving Safety Benchmark for Foundation Models (arxiv.org)
46
1
MarineEVT: Advancing Event-Centric Marine Video Understanding via Visual Tool Reasoning (arxiv.org)
47
1
CAPT: A Multi-task Continuous Autoregressive Transformer enabling Cross-dataset and Cross-species Transfer for Calcium Population Dynamics (arxiv.org)
48
1
MS-GPT: Rethinking MS/MS De Novo Structure Elucidation as Spectrum-Induced Posterior Querying of a Molecule-Language Model (arxiv.org)
49
1
Extreme Volatility Warning under Label Scarcity via Multi-Source Anomaly Fusion (arxiv.org)
50
1
Reason Popper-ly: Patching In-Context Reasoning with Inductive Logic Programming (arxiv.org)