TreeRU Tech Blog
Insights and hands-on guides on frontend, backend, databases, IT infrastructure, and AI solutions.
AI
Trending Top 5
Categories
posts.tsx — 28 posts
Load More18 remaining
- Local LLM Benchmark: 6 Models Tested Across 60 Questions and 7 Business ScenariosAI
- Building a Local RAG Pipeline — From Embedding Model Selection to Hallucination EliminationAI
- MoE vs Dense: Why Qwen3-30B-A3B Is Slower Than 14B — And Why Hybrid Inference FailedAI
- Qwen3-32B vs 14B — Is 2x Slower Speed Worth the Quality Gain?AI
- Text2SQL Real-World Test — When LLM Writes SQL DirectlyAI
- AWQ Quantization Speed Benchmark: 16 Models, INT4 vs BF16, and the MoE ReversalAI
- SGLang 23-Model Serving Guide — Optimal Configuration for Every ModelAI
- Qwen3-14B Deep Review — Why It Is Our Top-Ranked Local LLMAI
- Cross-Server AI Inference — Boosting Throughput 70% with a $450 Secondary GPUAI
- Serving 7 Companies on 1 GPU — Multi-Tenant Isolation Testing in PracticeAI
- LoRA Fine-Tuning for Custom AI Chatbots — From 10 Training Pairs to Multi-Tenant ServingAI
- SGLang vs vLLM — The Secret Behind the 3x Throughput GapAI
- Local LLM Concurrent User Load Test — How Many Users Can an RTX PRO 6000 Handle?AI
- Local LLM Business Test (Part 2) — Shopping, Legal & AutomationAI
- Local LLM Business Test (Part 1) — Manufacturing, SaaS, HealthcareAI
- LLM Hallucination Test — Which Local Models Fabricate Information?AI
- Local LLM Korean Language Comparison — 6 Models Tested with 10 Real QuestionsAI
- 8B vs 14B vs 32B LLM: Concurrent User Benchmark on a Single GPUAI
Editor's Picks