LiteLLM journey: drop_params ทำไมต้อง recreate
เรื่องจริงที่เกือบจะ apply drop_params=true ผ่าน DB อย่างเดียว แต่ error ยังไม่หายจนกว่าจะ recreate container — lesson เรื่อง DB auto-reload scope
เรื่องจริงที่เกือบจะ apply drop_params=true ผ่าน DB อย่างเดียว แต่ error ยังไม่หายจนกว่าจะ recreate container — lesson เรื่อง DB auto-reload scope
Investigation log: qwen CLI v0.23.x เปลี่ยน wire format → MiniMax reject ด้วย 400 (2013) → เลือกแก้ที่ gateway ทำไม ไม่ PR upstream ทำไม
Optimize Qwen3.8-Flash-Next (180B MoE) บน single DGX Spark — TTFT warm 5s → 1.4s, +7.4% throughput, ใช้แค่ 2 patches + 1 config tweak
บันทึกเรื่อง meme '1 year on Linux vs 10 years on GNU/Linux' แล้วสำรวจ setup จริงบนเครื่อง T14 เพื่อดูว่าผมเป็นสาย Linux ระดับไหน
วิเคราะห์โครงการ OpenAI x MHESI AI Accelerator ที่เพิ่งเปิดตัว 28 ส.ค. 2026 — ทำไม 8 สัปดาห์ + API credit $2,000 ไม่ใช่ infra และไทยควรลงทุนกับ model, education, และ critical understanding ของ AI แทน
เจาะลึก QSA, Gated Residual, N-gram Embedding, Muon+AdamW ของ Qwen3.8-Flash-Next จาก tech report 51 หน้า — ทั้งข้อดี ข้อเสีย และข้อถกเถียงจาก community
บทวิเคราะห์กลไกที่ธุรกิจสีเทาใช้ธุรกิจสีขาวเป็นฉากบังหน้า ฟอกเงิน สร้างความน่าเชื่อถือ และใช้ใบอนุญาตเป็นโล่ — พร้อมกรณีศึกษาคลินิกความงามไทย
5 เคสจริงที่ White ชนะ Gray — หวยรัฐ, Central Group, Wattanosoth, Agoda/Central Hotels, และกรมสรรพากร พร้อม pattern ของผู้ชนะ 5 ข้อ
7-step slippery slope — ทำไม White ถึงถูกบีบจนเลือกลดมาตรฐาน และจะป้องกันตัวเองได้ยังไง พร้อม 5 red flags
Qwen เปิดตัว Qwen3.8-Flash-Next บน Hugging Face 26 ส.ค. 2026 — 125B MoE + 51B n-gram embedding + 4B MTP, active แค่ 6B แต่ agentic coding/tool use ชนะ Claude Opus 4.6 Max; benchmark พร้อม NVFP4 quant จาก RadixArk ที่ลด 360 GB → 135 GB; วิเคราะห์ว่าลง DGX Spark GB10 128 GB ได้ไหม