LIVE · refreshes every 20 min
updated Sep 16, 05:43 UTC
110010
.art
(ificial intelligence)
AI products, research, and launches — clustered and ranked, not just re-blogged.
All stories
Research
Discussion
RSS
⌕
Search
/
☾
110010
.art
/ topic
Qwen
Sep 14
1d ago
Accelerating Dropless MoE training in JAX with NVIDIA Transformer Engine
NVIDIA Developer
→ story
Sep 11
4d ago
Qwen 0.8B shows strong multilingual understanding in long-form tasks
Hacker News (AI)
→ story
Sep 7
9d ago
What Does Multi-Harness RL Learn? Credit assignment and portability in coding agents
arXiv cs.AI
→ story
Sep 6
9d ago
Qwen-Scope: decoding intelligence, unleashing potential
Hacker News (AI)
→ story
Sep 4
11d ago
Show HN: Sageling — free, private, local, cowork AI harness for non-devs
Hacker News (AI)
→ story
12d ago
LLMs temper Bayesian priors using a single unembedding direction, the direction of ignorance
arXiv cs.LG
→ story
Sep 3
12d ago
Qwen 3.8 27B available on Cerebras at 1500 tok/sec
Hacker News (AI)
→ story
13d ago
Post-training ternarization of Qwen3-4B: effective bit budget, storage compression, and deployment
arXiv cs.AI
→ story
13d ago
Show HN: DeltaCode beats Qwen's zvec-grep on our test (92% vs 48%)
Hacker News (AI)
→ story
Sep 1
15d ago
The Halt Vector internalizes a causal steering intervention to halt chain-of-thought reasoning for more efficient reasoning
arXiv cs.LG
→ story
Aug 31
16d ago
Below the noise floor: bimodal seed collapse and distinct failure modes in small-model knowledge distillation
arXiv cs.CL
→ story
16d ago
First make it playable, then make it good: staged interaction learning for small dialogue-game agents
arXiv cs.CL
→ story
Aug 21
25d ago
Qwen image 3.0 Pro vs. GPT image 2 for production image APIs
Hacker News (AI)
→ story
Aug 19
27d ago
Ornith-1.5 9B may not be bad after all
r/LocalLLaMA
→ story
27d ago
QWEN plans to reuse the 397B-A17B architecture to compete with Deepseek V4 0731 Flash
r/LocalLLaMA
→ story
27d ago
NVFP4 on VOLTA matches RTX 5090 for Qwen 3.8 in FP4/FP8 on four 2017 Tesla V100s
r/LocalLLaMA
→ story
27d ago
Prompt extend model fine-tuned from gemma-4-12B-it for Qwen Image Edit 2511 generates enhanced editing prompts using PERL with ROLL; Kimi K2.6 serves as reward worker to evaluate results
r/LocalLLaMA
→ story
27d ago
llama.cpp adds --n-cpu-ffn option for Dense models (building on --n-cpu-moe / --cpu-moe for MOE models) via pull request 26622
r/LocalLLaMA
→ story
27d ago
Qwen 3.8 27B performs worse than Claude Code and GitHub Copilot for agentic coding, according to user experience with local models and tools
r/LocalLLaMA
→ story