LIVE · refreshes every 20 min
updated Sep 16, 03:43 UTC
110010
.art
(ificial intelligence)
AI products, research, and launches — clustered and ranked, not just re-blogged.
All stories
Research
Discussion
RSS
⌕
Search
/
☾
110010
.art
/ archive
Research
Sep 12
3d ago
air: The Agent Incident Registry (AIR) catalogs agent-related events with evidence and identifiers to aid comparisons with agent-security evaluations; it aims to prevent repeated AI agent failures
arXiv cs.AI
→ story
3d ago
An Open Recipe for IMO Gold: Training Nemotron for Olympiad Mathematics
arXiv cs.AI
→ story
3d ago
ARCHE: an autonomous agentic system for mechanistic discovery in chemistry using integrated reasoning, chemistry modeling, and tool registry
arXiv cs.AI
→ story
3d ago
Auditing update admission for continual embodied agents; independent evaluation can restrict useful continual learning while rejecting harmful policy updates.
arXiv cs.AI
→ story
3d ago
Automating QUBO formulation generation from natural language
arXiv cs.AI
→ story
3d ago
Debate-to-Skill: capability-bound process supervision for industrial query-to-agent annotation
arXiv cs.AI
→ story
3d ago
Decoupling readiness from release for tail-aware scheduling of agentic LLM workflows
arXiv cs.AI
→ story
3d ago
Defining AI agents: a compendium of criteria, metrics, and benchmarks
arXiv cs.AI
→ story
Sep 11
4d ago
Lifesaving AI-GUIDE device wins 2026 Excellence in Technology Transfer Award
MIT News (AI)
→ story
4d ago
More components do not always improve P300 speller performance; a four-component full-factorial test finds anti-synergy among Euclidean Alignment, xDAWN filtering, calibration, and language model priors.
arXiv cs.LG
→ story
4d ago
Motif-oriented graph captioning is explored as bidirectional graph-text translation to abstract connectivity into recognizable motifs.
arXiv cs.CL
→ story
4d ago
Multi-Agent Agentic Graph Learning via Structural Signatures
arXiv cs.AI
→ story
4d ago
Multilingual in name only? Cultural and linguistic weaknesses of LLMs in Urdu
arXiv cs.CL
→ story
4d ago
NCP-ArchPreview advances latent-space language modeling with Next Concept Prediction alongside next-token prediction
arXiv cs.CL
→ story
4d ago
OpenDiscoveryTrace provides a public dataset of 558 complete AI scientific agent trajectories for evaluating AI scientist workflows
arXiv cs.AI
→ story
4d ago
Overview of the NLPCC 2026 shared task 11: agent-based experiment reproduction from scientific papers
arXiv cs.CL
→ story
4d ago
Phases in a class of associative memories via hidden neurons
arXiv cs.LG
→ story
4d ago
PRAGMA: evaluating personalized guidance with memory alignment in lifelong conversations
arXiv cs.AI
→ story
4d ago
Processing and classifying invasive bird species vocalizations in noisy natural soundscapes using Bayesian wavelet shrinkage and supervised learning
arXiv cs.LG
→ story
4d ago
ProMediConv benchmarks proactive conversational agents in legal dispute mediation; arXiv paper proposes a multi-stage mediation framework
arXiv cs.CL
→ story
4d ago
Rebalancing token importance in language models with TF-IDF weighted cross-entropy loss
arXiv cs.CL
→ story
4d ago
Relatively Smart II: tractable or semi-supervised instance-optimal learning
arXiv cs.LG
→ story
4d ago
RESCUE-BENCH: relation-aware multi-party emotional support conversation systems
arXiv cs.AI
→ story
4d ago
Risk-constrained stopping layer Cros enables autonomous diagnosis timing in sequential clinical agents; evaluates state-wise error and selective diagnostic errors using LTT-style tests.
arXiv cs.AI
→ story
4d ago
RiVaT-Fuse: reliability-calibrated variational tensor fusion for multimodal prediction under modality uncertainty
arXiv cs.LG
→ story
4d ago
Robust multimodal sentiment analysis with incomplete modalities via semantic-aware completeness-based reconstruction
arXiv cs.CL
→ story
4d ago
RobustSGPO introduces controlled search-space edits for agent harness evolution; evaluates permission scheduling, cumulative controls, and task-family transfer in AgentX brains
arXiv cs.AI
→ story
4d ago
Rubric-aligned disentangled evaluation of human simultaneous interpreting
arXiv cs.CL
→ story
4d ago
SearchAtlas analyzes agentic search strategies by converting search trajectories into evidential graphs
arXiv cs.CL
→ story
4d ago
Seven sources of physical AI capability formation
arXiv cs.AI
→ story
4d ago
The menu is an execution prior: state-path tool menus for online agents
arXiv cs.AI
→ story
4d ago
The Truth Was Never Gone: Perfect Aliasing in Compliant-Context Truth Probes
arXiv cs.LG
→ story
4d ago
Think Before You Link: rarity, reasoning, and retrieval in multilingual entity linking
arXiv cs.CL
→ story
4d ago
Using semantic uncertainty to estimate transition relevance in turn-taking
arXiv cs.CL
→ story
4d ago
Valerant: an automatic navigable game map generator via action-conditioned world model exploration
arXiv cs.AI
→ story
4d ago
Verbalized confidence becomes a more robust soft-scoring signal than log-probabilities for LLM-as-a-Judge on post-2025 proprietary models; the authors term this a compatibility shift.
arXiv cs.CL
→ story
4d ago
When Noise Fabricates Bias: the fragility of LLM-as-a-judge bias measurement under noisy text
arXiv cs.CL
→ story
4d ago
Which tokens should SFT actually learn? A token-trimming perspective on mathematical reasoning
arXiv cs.AI
→ story
4d ago
XAI-Arena: LLMs assess the quality of XAI explanations as a scalable judge; arXiv paper explores using LLMs as reproducible evaluations
arXiv cs.AI
→ story
4d ago
Zero-shot rib design: merging training-free generative prior with topology optimization
arXiv cs.LG
→ story
←
1
2
3
4
5
6
7
8
9
10
11
12
13
14
15
16
17
18
19
20
21
22
23
24
25
26
27
28
29
30
31
32
33
34
→
Older items: monthly archive →