LIVE · refreshes every 20 min
updated Sep 16, 03:43 UTC
110010
.art
(ificial intelligence)
AI products, research, and launches — clustered and ranked, not just re-blogged.
All stories
Research
Discussion
RSS
⌕
Search
/
☾
110010
.art
/ archive
Research
Sep 5
10d ago
A computable representation of the physical laboratory enables verifiable workflows
arXiv cs.AI
→ story
10d ago
Aggressive decoding-time KV eviction: EMA temporal aggregation preserves ranking under aggressive KV compression
arXiv cs.AI
→ story
10d ago
AI-driven practical English textbooks: five-layer architecture for adaptive learning and feedback
arXiv cs.AI
→ story
10d ago
Analysis of prompt engineering for drug toxicity prediction
arXiv cs.AI
→ story
10d ago
Attention triangle in audio-video diffusion models reveals cross-modal semantic leakage
arXiv cs.AI
→ story
10d ago
AutoGraphForge advances automated graph-theoretic conjecturing, refuting, formalizing, and proving via a counterexample-guided pipeline
arXiv cs.AI
→ story
10d ago
Beyond “Made with AI”: visualizing provenance density to mitigate the transparency penalty
arXiv cs.AI
→ story
10d ago
Caught in the story: narrative captivity in multi-turn LLM conversations
arXiv cs.AI
→ story
10d ago
CulturalMenuBench probes the knowledge-application gap in multimodal culinary reasoning across 10 languages and 18 regions with 4,870 items and 10 tasks
arXiv cs.AI
→ story
10d ago
Dalek: a constructive agent machine that self-maintains, self-evolves, self-reproduces, and self-organizes on any suitable substrate
arXiv cs.AI
→ story
10d ago
Do GUI agents know when not to act? Enabling conflict-aware termination for multimodal GUI agents
arXiv cs.AI
→ story
10d ago
Dude: A dual-detection multi-agent system for paper-code discrepancy detection
arXiv cs.AI
→ story
10d ago
DuplexSpeechBench-IFEval evaluates implicit instruction following in full-duplex voice agents.
arXiv cs.AI
→ story
10d ago
Feature reconfiguration with visual prior for medical lesion segmentation
arXiv cs.AI
→ story
10d ago
GPS-Bench: a governance policy benchmark for automating policy analysis
arXiv cs.AI
→ story
10d ago
GrowPage: on-demand KV budgeting for efficient LLM reasoning serving
arXiv cs.AI
→ story
10d ago
HalluPeer: a taxonomy-driven benchmark for detecting hallucinations in scientific peer reviews
arXiv cs.AI
→ story
10d ago
Making every tool call count: necessary tool-evidence path rewards for agentic vision-language models
arXiv cs.AI
→ story
Sep 4
11d ago
LLMs’ use of memory and user context biases financial analysis, study finds across 3,575 SEC filings
arXiv cs.CL
→ story
11d ago
Margins, not windows: training-free per-step lossy speculative decoding
arXiv cs.CL
→ story
11d ago
Matrix-CODI on ProsQA shows rank-indifference in latent matrices and a failure of low-rank truncation to hurt accuracy
arXiv cs.LG
→ story
11d ago
MemoryLACE: memory lifecycle–aware consolidation and evidence retrieval
arXiv cs.CL
→ story
11d ago
Mesh-native physics-informed graph surrogates for TCAD-in-the-loop design space exploration
arXiv cs.LG
→ story
11d ago
Modern Transformers are implicit hybrids: designing principled hybrid architectures by analyzing head-level functional organization in RoPE-based transformers
arXiv cs.LG
→ story
11d ago
No country for old linguists: LLM-brain alignment underdetermines neural computation
arXiv cs.CL
→ story
11d ago
No-Regret Bayesian optimization with finite-library input-warped kernels
arXiv cs.LG
→ story
11d ago
ObserverBench benchmarks whether an internal observer can reliably guide intervention and control tasks
arXiv cs.LG
→ story
11d ago
PiPMRE: a pipeline based on language model for medical relation extraction
arXiv cs.CL
→ story
11d ago
Portable causal fairness across synthetic data generator families
arXiv cs.LG
→ story
11d ago
R^2Adapter: a routing and rewriting adapter for efficient hybrid RAG
arXiv cs.CL
→ story
11d ago
RL-ADA: a world-feedback framework for adversarially robust enterprise dialogue agents
arXiv cs.CL
→ story
11d ago
Scaling laws, tabular data, and actuarial ratemaking models are explored in a real-world motor insurance portfolio
arXiv cs.LG
→ story
11d ago
Self-correcting speech recognition in large audio language models through hidden-state interactions; latent-state-based refinement improves warm-initialized LLM ASR using base LLMs and LoRA adaptation
arXiv cs.CL
→ story
11d ago
SHELF: a synthetic harness for evaluating LLM fitness in multi-task bibliographic benchmarking
arXiv cs.CL
→ story
11d ago
SWIM: Student Writing Simulation via Proficiency-Conditioned Generation
arXiv cs.CL
→ story
11d ago
Tail-Likelihood Reinforcement Learning; researchers propose optimizing coverage to distinguish policies with identical mean rewards but different rare high-reward rollout probabilities.
arXiv cs.LG
→ story
11d ago
The 2026 PNPL competition achieves word classification and cross-subject generalisation in LibriBrain100
arXiv cs.LG
→ story
11d ago
TRACE: spatiotemporal contact memory graph network simulator for granular dynamics
arXiv cs.LG
→ story
11d ago
Unifying conformal language tasks with in-context ensembles
arXiv cs.CL
→ story
11d ago
Using counterexamples as feedback for agent self-correction in NL-to-regex synthesis with A-CEGIS
arXiv cs.CL
→ story
←
1
2
3
4
5
6
7
8
9
10
11
12
13
14
15
16
17
18
19
20
21
22
23
24
25
26
27
28
29
30
31
32
33
34
→
Older items: monthly archive →