DEV Community

#llm

Posts

đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.
Amazon Nova 2: A Developer's Guide to Lite, Pro, and Omni

Amazon Nova 2: A Developer's Guide to Lite, Pro, and Omni

Comments 1
7 min read
Battle-Tested Multi-Agent Orchestration Patterns with Google ADK: Parallel, Sequential, and Persistent Sessions

Battle-Tested Multi-Agent Orchestration Patterns with Google ADK: Parallel, Sequential, and Persistent Sessions

Comments 1
4 min read
Claude 3.5 Sonnet is more than an upgrade. It’s a new workflow.

Claude 3.5 Sonnet is more than an upgrade. It’s a new workflow.

Comments
3 min read
Why I chose Gemma4b over Mistral 7b?

Why I chose Gemma4b over Mistral 7b?

Comments
4 min read
What Model Quantization Actually Does: From Float16 to 4-Bit Weights

What Model Quantization Actually Does: From Float16 to 4-Bit Weights

1
Comments 3
6 min read
What Happens When an AI Agent Gets Stuck in a Loop?

What Happens When an AI Agent Gets Stuck in a Loop?

5
Comments
7 min read
DeepSeek V4.1 Flash API Cost: Matches V4 Pro at 3-6x Less per Answer

DeepSeek V4.1 Flash API Cost: Matches V4 Pro at 3-6x Less per Answer

Comments
12 min read
Letting an LLM write alerting rules, but never letting it flip the switch

Letting an LLM write alerting rules, but never letting it flip the switch

Comments 3
5 min read
llama.cpp vs Ollama in 2026: Which Runtime Should You Run?

llama.cpp vs Ollama in 2026: Which Runtime Should You Run?

Comments
18 min read
Implementing AI Observability: Tracing LLM Calls End-to-End

Implementing AI Observability: Tracing LLM Calls End-to-End

Comments
4 min read
Your model confabulates

Your model confabulates

Comments
6 min read
How I Fine-Tuned a 7B LLM with LoRA and Unsloth

How I Fine-Tuned a 7B LLM with LoRA and Unsloth

Comments
3 min read
How I Manage API Keys for 4 Providers Behind One Client

How I Manage API Keys for 4 Providers Behind One Client

Comments
6 min read
Building AI Observability for the Native Stack: Architecture Design and Engineering Practice from Bonree ONE 4.0

Building AI Observability for the Native Stack: Architecture Design and Engineering Practice from Bonree ONE 4.0

Comments
14 min read
OpenArch: PyTorch implementations of modern LLM architectures!

OpenArch: PyTorch implementations of modern LLM architectures!

Comments
5 min read
đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.