skip to main content
📰
NewsNook
nesting hacker news in a more meaningful way
≡
menu
(⌘/)
LLM / page 50
newer
older
your nooks
add yours
Nobody owns the 30 seconds an LLM spends thinking, so I made an arcade there
Engy – Verified LLM Inference
Show HN: Wingman – A client-agnostic agent harness (written in Go)
Show HN: VernLLM – The AI resilience layer for TypeScript
Pushing the limits of RISC-V emulation
Show HN: MandoCode Desktop – Native Windows AI Coding Assistant on Ollama .NET
KaaS – Knowledge as a Service: an out-of-the-box LLM wiki compiler
Show HN: Energy, carbon and water estimates for AI content, shown as ranges
CryptanalysisBench: Can LLMs Do Cryptanalysis?
Launch HN: Tokenless (YC S26) – Automatic model switching to save money
What NPU Tops Doesn't Tell You: How Memory Bounds Local LLM Performance
The Scientific Literature Is Poisonous to LLMs
Show HN: Open-source engine running Gemma 4 26B in 2 GB RAM on any M-series Mac
Sparse Attention with Persistent State Machines – High‑Sparsity LLM Accelerator
Escaping the LLM Coding Rat Race
Show HN: NightRun, bare metal LLM inference, no OS, boots from USB
Codex's 5-hour usage limit returns tomorrow
Show HN: Multi-agent LLM editor with local inference via WebSockets