Skip to content
Navigation menu
Search
Powered by Algolia
Search
Log in
Create account
DEV Community
Close
#
llm
Follow
Hide
Posts
Left menu
👋
Sign in
for the ability to sort posts by
relevant
,
latest
, or
top
.
Right menu
llama.cpp vs Ollama — which one should you run?
Mr Say Nothing
Mr Say Nothing
Mr Say Nothing
Follow
Oct 6
llama.cpp vs Ollama — which one should you run?
#
discuss
#
ai
#
ollama
#
llm
5
reactions
Comments
1
comment
5 min read
Why Byte-Faithful Pass-Through Beats Protocol Translation for Tool-Calling Agents
DHPP
DHPP
DHPP
Follow
Oct 6
Why Byte-Faithful Pass-Through Beats Protocol Translation for Tool-Calling Agents
#
ai
#
agents
#
llm
#
router
Comments
3
comments
4 min read
MCP Connected Your Tools. It Didn't Fix Your Agent's Memory.
Shweta Mishra
Shweta Mishra
Shweta Mishra
Follow
Oct 6
MCP Connected Your Tools. It Didn't Fix Your Agent's Memory.
#
ai
#
programming
#
llm
#
mcp
3
reactions
Comments
2
comments
4 min read
Benchmark Contamination 101: How Train/Test Overlap Inflates Leaderboard Scores (and How to Catch It)
Ward Ed
Ward Ed
Ward Ed
Follow
Oct 6
Benchmark Contamination 101: How Train/Test Overlap Inflates Leaderboard Scores (and How to Catch It)
#
machinelearning
#
llm
#
benchmark
#
datascience
Comments
Add Comment
7 min read
KV Cache Quantization in LLM Serving: FP8 and INT8 Tradeoffs, the Silent config.json Trap, and How to Measure It Fairly
AI Tech News
AI Tech News
AI Tech News
Follow
Oct 6
KV Cache Quantization in LLM Serving: FP8 and INT8 Tradeoffs, the Silent config.json Trap, and How to Measure It Fairly
#
llm
#
vllm
#
performance
#
devops
Comments
Add Comment
7 min read
Restoring a grant is not restoring capacity
jaycodes
jaycodes
jaycodes
Follow
Oct 6
Restoring a grant is not restoring capacity
#
llm
#
ai
#
performance
#
mlops
Comments
Add Comment
9 min read
When the attacker is an agent: a defender's field guide to autonomous AI intrusions in 2026
ai maya
ai maya
ai maya
Follow
Oct 6
When the attacker is an agent: a defender's field guide to autonomous AI intrusions in 2026
#
ai
#
security
#
llm
#
machinelearning
Comments
Add Comment
6 min read
OpenAI Started Watermarking ChatGPT Text. Build a Tiny Text Watermark in TypeScript.
Bobby Hall Jr
Bobby Hall Jr
Bobby Hall Jr
Follow
Oct 6
OpenAI Started Watermarking ChatGPT Text. Build a Tiny Text Watermark in TypeScript.
#
ai
#
typescript
#
llm
#
machinelearning
2
reactions
Comments
1
comment
12 min read
Do LLMs Invent Japanese Law Articles? A Bilingual Benchmark
Raihan
Raihan
Raihan
Follow
Oct 6
Do LLMs Invent Japanese Law Articles? A Bilingual Benchmark
#
llm
#
nlp
#
japanese
#
opensource
Comments
Add Comment
4 min read
How to Build Resilient AI Agents with Search Fallback Loops
Pratik
Pratik
Pratik
Follow
Oct 6
How to Build Resilient AI Agents with Search Fallback Loops
#
ai
#
llm
#
architecture
#
programming
Comments
Add Comment
5 min read
4-bit GGUF Quality for MoE Models: Why Only 3B of 180B Params Fire, and How to Prove Parity
GINIGEN AI
GINIGEN AI
GINIGEN AI
Follow
Oct 6
4-bit GGUF Quality for MoE Models: Why Only 3B of 180B Params Fire, and How to Prove Parity
#
ai
#
llm
#
quantization
#
edgeai
Comments
Add Comment
7 min read
Indirect Prompt Injection Through Tool Descriptions and Tool Output: How Untrusted Metadata Hijacks Agents
hugginf_expert
hugginf_expert
hugginf_expert
Follow
Oct 6
Indirect Prompt Injection Through Tool Descriptions and Tool Output: How Untrusted Metadata Hijacks Agents
#
ai
#
agents
#
security
#
llm
Comments
Add Comment
7 min read
Le Chonk Unleashed: Why Mistral’s 1‑Trillion‑Parameter Model Is Shattering LLM Benchmarks
amrit
amrit
amrit
Follow
Oct 6
Le Chonk Unleashed: Why Mistral’s 1‑Trillion‑Parameter Model Is Shattering LLM Benchmarks
#
mistral
#
llm
#
openai
#
ai
Comments
Add Comment
5 min read
Claude Opus 5.5: What Changed, What Breaks, and Whether to Switch
Nancy Garg
Nancy Garg
Nancy Garg
Follow
for
Studio1
Oct 6
Claude Opus 5.5: What Changed, What Breaks, and Whether to Switch
#
ai
#
claude
#
llm
#
programming
Comments
Add Comment
7 min read
ZTC (Zero-Token Confidence): juzgar una acción de IA leyendo el estado interno del modelo, con cero tokens extra
김민식/학생
김민식/학생
김민식/학생
Follow
Oct 6
ZTC (Zero-Token Confidence): juzgar una acción de IA leyendo el estado interno del modelo, con cero tokens extra
#
ai
#
llm
#
security
#
spanish
Comments
Add Comment
6 min read
👋
Sign in
for the ability to sort posts by
relevant
,
latest
, or
top
.
We're a place where coders share, stay up-to-date and grow their careers.
Log in
Create account