LIVE · refreshes every 20 min
updated Sep 16, 03:43 UTC
110010
.art
(ificial intelligence)
AI products, research, and launches — clustered and ranked, not just re-blogged.
All stories
Research
Discussion
RSS
⌕
Search
/
☾
110010
.art
/ archive
Research
Sep 2
13d ago
Adapting without gradients: affine statistics transport and what its certificate can tell you
arXiv cs.LG
→ story
13d ago
AI Morbidity and Mortality: a framework for clinical AI failure review
arXiv cs.AI
→ story
13d ago
AI should not only be helpful. It should be contingent; artificial intimacy, sycophancy, and the future of social learning
arXiv cs.AI
→ story
13d ago
Assessing alignment and stability of feature importance explanations via weight of evidence
arXiv cs.LG
→ story
Sep 1
14d ago
Walter Torous named executive director of MIT Center for Real Estate
MIT News (AI)
→ story
14d ago
Mapping global methane emissions from space with deep learning
Google Research
→ story
14d ago
PCFBench: a diagnostic benchmark for product carbon footprint estimation
arXiv cs.AI
→ story
14d ago
Preference elicitation over allocation outcomes for policy optimization to align heart transplantation with human values, via learning a utility function from stakeholder preferences
arXiv cs.AI
→ story
14d ago
Probing perceptual priors of MLLMs via Gibbs sampling with interpretable generative controls
arXiv cs.AI
→ story
14d ago
PromptKWS: a prompt-guided open-vocabulary keyword spotting framework introduces the Prompt Phrases Prediction Network for embedding keyword prompts
arXiv cs.CL
→ story
14d ago
Race between agentic AI capabilities and data quality control in online surveys
arXiv cs.AI
→ story
14d ago
RankShift detects and explains in-database shifts in category shares without changing overall event counts
arXiv cs.LG
→ story
14d ago
Rating the raters: Rasch measurement theory for LLM evaluation
arXiv cs.AI
→ story
14d ago
RegDivergence-101: an LLM benchmark for cross-jurisdiction regulatory contradiction detection in life sciences
arXiv cs.AI
→ story
14d ago
ReToolSQL: reinforcement learning from execution feedback with iterative refinement for text-to-SQL
arXiv cs.AI
→ story
14d ago
Retrieving relations, detecting fallacies: a rag approach to political debate analysis
arXiv cs.AI
→ story
14d ago
ReVA: a region-aware visual assistant for visually grounded question answering
arXiv cs.CL
→ story
14d ago
Revisiting the provable-auditable privacy gap of DP-SGD
arXiv cs.LG
→ story
14d ago
Reward-Oracle MCTS for formal theorem proving: sample-efficient search and the need for kernel-level proof auditing
arXiv cs.AI
→ story
14d ago
Rigour-matched audit compares periodic-step layer skipping methods ConfLayers and SWIFT for efficient LLM inference; includes analysis of trained routing alternatives
arXiv cs.CL
→ story
14d ago
Self-evolving skills via surrogate-guided solve-and-reproduce
arXiv cs.AI
→ story
14d ago
SETU: an agentic ecosystem for multilingual, persona-aware communication coaching
arXiv cs.AI
→ story
14d ago
SHAPE analyzes chain-of-thought trajectories in math reasoning using semantic spaces and steps through mathematical interpretations
arXiv cs.AI
→ story
14d ago
Sparse Koopman autoencoders identify local dynamical regimes in multibasin systems
arXiv cs.LG
→ story
14d ago
STAGEET: Stage-wise Typed Edit Tagging for grammatical error correction with Arabic as a case study
arXiv cs.CL
→ story
14d ago
Statutory AI: aligning large language models with legal norms
arXiv cs.AI
→ story
14d ago
Temperature-Adaptive Transformed Teacher Matching
arXiv cs.LG
→ story
14d ago
Terminal-Bench-LILT proposes a multilingual coding benchmark with 300 tasks across ten languages to evaluate non-English software development challenges.
arXiv cs.CL
→ story
14d ago
Test-time scaling improves LLM-driven automated equation discovery by enabling iterative search with additional compute at test time.
arXiv cs.CL
→ story
14d ago
The Halt Vector internalizes a causal steering intervention to halt chain-of-thought reasoning for more efficient reasoning
arXiv cs.LG
→ story
14d ago
The signal in the noise: an auditable reliability layer for biomedical text classification
arXiv cs.AI
→ story
14d ago
Thinking costs tokens: adding inference structure hurts performance below a token-budget threshold and helps above it
arXiv cs.AI
→ story
14d ago
Time Capsule of testable human knowledge: 41 years of Jeopardy! in a single free local model
arXiv cs.AI
→ story
14d ago
Titans-QFWP: a regime-aware hybrid quantum fast weight programmer for portfolio optimization
arXiv cs.LG
→ story
14d ago
TPvG: A Moral Decision Framework for Large Language Models from One-Shot to Sequential Feedback
arXiv cs.AI
→ story
14d ago
Unsupervised latent space alignment with hyperspherical geodesic matching
arXiv cs.LG
→ story
14d ago
Using generative AI to design accessible interactive visualizations for undergraduate mathematics via a six-phase workflow (Foundation, Customization, Mathematical Depth, Application, Accessibility, Pedagogical Control)
arXiv cs.AI
→ story
14d ago
V2TATC: a joint voice-trajectory embedding framework and dataset for air traffic controller situational awareness
arXiv cs.LG
→ story
14d ago
Why Didn’t It Check? Unsupported final claims and their repair in two tool-equipped language models
arXiv cs.AI
→ story
14d ago
WM-R1 trains mobile GUI agents with world models instead of real environments for reinforcement learning
arXiv cs.AI
→ story
←
1
2
3
4
5
6
7
8
9
10
11
12
13
14
15
16
17
18
19
20
21
22
23
24
25
26
27
28
29
30
31
32
33
34
→
Older items: monthly archive →