LIVE · refreshes every 20 min
updated Sep 16, 03:23 UTC
110010
.art
(ificial intelligence)
AI products, research, and launches — clustered and ranked, not just re-blogged.
All stories
Research
Discussion
RSS
⌕
Search
/
☾
110010
.art
/ archive
Research
Sep 10
5d ago
A statistical approach to estimating sample size of machine learning models
arXiv cs.LG
→ story
5d ago
Accountable and uncertainty-aware evaluation of sensor-based AI under distribution shift: devices, subjects, and nearly three years underground
arXiv cs.LG
→ story
5d ago
Applying foundation model embeddings to assess urban livability using high-resolution geospatial data
arXiv cs.LG
→ story
5d ago
Auditable emergency triage for maternal and newborn care in India; Noora Health uses an LLM to classify emergencies in WhatsApp messages for on-demand caregiver support
arXiv cs.CL
→ story
5d ago
Auditors fabricate: batch-size degradation and confident hallucination in LLM detection of planted document contamination
arXiv cs.CL
→ story
5d ago
Benchmarking hybrid deep research across database querying and web search
arXiv cs.CL
→ story
5d ago
BuzzASR: a swarm of 100+ language-specialized Whisper ASR models for 102 languages
arXiv cs.CL
→ story
5d ago
CARRE: Counterfactual Action Retrieval and Reason Evaluation for explainable churn prescription
arXiv cs.CL
→ story
5d ago
Constraint-aware discrete black-box optimization using tensor decomposition
arXiv cs.LG
→ story
5d ago
DiffLUT-Net: differentiable training of FPGA LUT networks with learnable connectivity
arXiv cs.LG
→ story
Sep 9
6d ago
MIT Schwarzman College of Computing launches pilot to help educators teach AI across disciplines
MIT News (AI)
→ story
6d ago
Newton Matching: a unified framework for fine-tuning and sampling in generative modeling
arXiv cs.LG
→ story
6d ago
Nonlinear elliptic homogenization with the parametric Deep Ritz method
arXiv cs.LG
→ story
6d ago
Online learning with LLM experts from limited feedback
arXiv cs.LG
→ story
6d ago
PAC-Private autoregressive generation calibrates noise to ensemble disagreement to protect private text predictions
arXiv cs.LG
→ story
6d ago
PGP-Clinical-TimeKAN jointly forecasts multivariate clinical trajectories with a prior-guided probabilistic approach
arXiv cs.AI
→ story
6d ago
Planning and scheduling business processes under control-flow uncertainty
arXiv cs.AI
→ story
6d ago
RAPID: reliability-aware pair importance distillation reduces quadratic costs by separating a reliability-gated relational target from a full support adaptive pair proposal.
arXiv cs.AI
→ story
6d ago
Reasoning-aware compression benchmarks per-module quantization across five reasoning benchmarks to protect vulnerable circuits in energy-efficient LLM deployment
arXiv cs.AI
→ story
6d ago
Recall is not protection: evaluating safety monitors against model compliance
arXiv cs.CL
→ story
6d ago
Recovering temporal and geographic signals from language model embeddings
arXiv cs.AI
→ story
6d ago
Robustness of LLM-generated SystemVerilog Assertions to semantics-preserving RTL transformations
arXiv cs.LG
→ story
6d ago
Rubric-guided large language model identifies opioid use disorder from EHRs using OPRO prompting
arXiv cs.CL
→ story
6d ago
SAFEGuard detects optimization-based jailbreak attacks through harmful semantic analysis and fluency measurement
arXiv cs.LG
→ story
6d ago
SCAFFOLD: Self-Improving Web Agents via Recursive Parametric Skill Abstraction
arXiv cs.AI
→ story
6d ago
Scaling optimal classification trees via adaptive feature and sample reduction via weighted STreeD reduces sample-dependent computation in fixed-candidate optimization
arXiv cs.LG
→ story
6d ago
SciLitBench: benchmark and design principles for LLM-powered systematic literature reviews
arXiv cs.AI
→ story
6d ago
Selective Posterior Margin Regularization for forward-corrected classification
arXiv cs.LG
→ story
6d ago
SinoGlyphBench: a diagnostic benchmark for Chinese glyph-level obfuscation in language-model moderation
arXiv cs.CL
→ story
6d ago
Solving versus verifying: catching contradictions in tax reasoning systems
arXiv cs.CL
→ story
6d ago
Some tokens act like magnets, revealing linguistic organization in language model layers
arXiv cs.CL
→ story
6d ago
SurveyAgent-HKA: a multi-agent framework for scientific survey generation with LLMs and human knowledge augmentation
arXiv cs.CL
→ story
6d ago
TamilEOT: a dataset and model for semantic end-of-turn detection in Tamil telephone speech
arXiv cs.CL
→ story
6d ago
The Failure Happens Before the Drift: The Social Dynamics of Values in LLM Agent Societies
arXiv cs.AI
→ story
6d ago
TopoBox-3D analyzes topology changes in neural PDE operators using Hodge heat flow; unseen domain topology alters invariant and decaying subspaces across six architectures
arXiv cs.LG
→ story
6d ago
UniRRM: unified reasoning reward models across languages and evaluation paradigms
arXiv cs.CL
→ story
6d ago
What if LLMs ate their words: causal history effects in multi-turn interaction
arXiv cs.CL
→ story
6d ago
What LLM trading agents actually do in production: a six-month, population-scale record from two fleets
arXiv cs.AI
→ story
6d ago
When do options help? policy necrosis and redundant coverage in option-critic
arXiv cs.LG
→ story
6d ago
Who Maintains Agent Skills? A longitudinal study of human-governed, AI-assisted skill maintenance
arXiv cs.CL
→ story
←
1
2
3
4
5
6
7
8
9
10
11
12
13
14
15
16
17
18
19
20
21
22
23
24
25
26
27
28
29
30
31
32
33
34
→
Older items: monthly archive →