LIVE · refreshes every 20 min
updated Sep 16, 03:23 UTC
110010
.art
(ificial intelligence)
AI products, research, and launches — clustered and ranked, not just re-blogged.
All stories
Research
Discussion
RSS
⌕
Search
/
☾
110010
.art
/ archive
Research
Aug 21
25d ago
FinSkillBench evaluates AI agents on financial domain skills for investment management.
arXiv cs.AI
→ story
25d ago
FleetSieve selects measurements for SLO-aware LLM fleet configuration to reduce unnecessary profiling; it targets resource-coupled decisions on tensor-parallel and replica counts
arXiv cs.LG
→ story
25d ago
Forking fast: efficiently estimating uncertainty dynamics in text generation
arXiv cs.CL
→ story
25d ago
FraudBench tests policy-grounded banking agents against adaptive fraud via stress-testing with customer data and authorization tasks
arXiv cs.AI
→ story
25d ago
Generating diverse personas for user simulators to test interview dialogue systems
arXiv cs.CL
→ story
25d ago
GenEx: a graph-based representational paradigm for SARS-CoV-2 variant detection via codon co-occurrence networks
arXiv cs.AI
→ story
25d ago
Hear2Act benchmarks when prosody should change what an assistant does.
arXiv cs.CL
→ story
25d ago
Holtercare-Bench: a multimodal benchmark for evaluating long-term dynamic ECG analysis
arXiv cs.LG
→ story
25d ago
Improved confidence estimates for black-box large language models
arXiv cs.LG
→ story
25d ago
Improving rural medication safety with AI: a scoping review
arXiv cs.AI
→ story
25d ago
Irrelevant text biases multimodal LLMs in visual tasks, study finds
arXiv cs.CL
→ story
25d ago
Jaccard distance satisfies the triangle inequality on arbitrary lattices for strictly positive, monotone, and modular valuations
arXiv cs.AI
→ story
25d ago
Kähler landscapes for complex neural network descents and guarantees including a search and destroy of the Calabi-Yau manifold
arXiv cs.LG
→ story
25d ago
Linguistic holonomy and statistical watermarks: inner geometry of meaning-preserving transformations
arXiv cs.CL
→ story
25d ago
LLM-Detector uses in-context learning for tabular anomaly detection
arXiv cs.LG
→ story
25d ago
Longitudinal Bayesian learning of continuous disease position across the Alzheimer's disease continuum
arXiv cs.LG
→ story
25d ago
MAS failures stem from concurrency control issues in LLM-based systems; long inference windows amplify stale reads and lost updates
arXiv cs.AI
→ story
25d ago
Mechanistic tomography: designed measurement for recovering control-oriented interpretability
arXiv cs.LG
→ story
25d ago
Metamorphic artificial age score decision-support prototype for flight-log-based drone propeller health monitoring developed using 2024 DronePropA flight logs.
arXiv cs.AI
→ story
25d ago
Mitigating identity essentialism in LLM agents with longitudinal life trajectories
arXiv cs.CL
→ story
25d ago
Mizo ASR system fine-tuned with three Whisper models and SraVaani 1.0; morphology-aware evaluation yields 7.22% WER at best
arXiv cs.CL
→ story
25d ago
NepOOC-M: Nepali-English benchmark and multimodal architectures comparison for out-of-context detection
arXiv cs.CL
→ story
25d ago
On-board implementation of an ML-based helicopter weight estimator using Airbus takeoff data and a learning assurance process per EASA and Eurocae ED-324
arXiv cs.LG
→ story
25d ago
Open-weight foundation models require better downstream governance than current model cards, study finds
arXiv cs.AI
→ story
25d ago
Optimized fuzzy logic approach with the IEEE Key Gas Method for diagnosing power transformer faults using dissolved gas analysis
arXiv cs.AI
→ story
25d ago
Profiling game worlds by transition complexity
arXiv cs.AI
→ story
25d ago
Quantifying event impacts on time series via multiscale contrastive learning
arXiv cs.LG
→ story
25d ago
Quantum kernel estimation for the discovery of early lung cancer detection
arXiv cs.LG
→ story
25d ago
Rationally enriched Chebyshev trunk bases for DeepONet surrogates of high-Péclet transport
arXiv cs.LG
→ story
25d ago
RDFdL: integrating RDF with Differential Dynamic Logic
arXiv cs.AI
→ story
25d ago
ReCache enables independent caching of resource representations to reduce inference-time overhead for tool-augmented LLM agents.
arXiv cs.CL
→ story
25d ago
Redakto introduces incognito tab for LLMs to protect privacy in EU contexts
arXiv cs.AI
→ story
25d ago
Reliable financial named entity recognition under domain shift shows confidence estimation and selective prediction across SEC filings, financial news, and general-topic text
arXiv cs.CL
→ story
25d ago
Remember, verify, or ask? Cross-family evaluation of memory commitment in LLM agents
arXiv cs.CL
→ story
25d ago
Represented but ignored: a causal account of prosodic underuse in audio-language models
arXiv cs.CL
→ story
25d ago
Self-evolving agents as dynamic graph transformation: a survey and new perspective
arXiv cs.AI
→ story
25d ago
Solving is not drawing: a benchmark for diagrammatic reasoning in Olympiad geometry
arXiv cs.AI
→ story
25d ago
SynFlow: a multidimensional diachronic semantic analysis toolkit
arXiv cs.CL
→ story
25d ago
The asymmetric harms of LLM compression reduce head knowledge retention more than overall knowledge across 3 models and 11 compression methods.
arXiv cs.CL
→ story
25d ago
Time-series retrieval grounds multimodal language models for remaining useful life estimation
arXiv cs.CL
→ story
←
1
2
3
4
5
6
7
8
9
10
11
12
13
14
15
16
17
18
19
20
21
22
23
24
25
26
27
28
29
30
31
32
33
34
→
Older items: monthly archive →