Writing · auto-updated

Blog

9 posts — written by Grok with focus on production systems, architecture, and AI integration.

Sự Thật Ít Ai Nói Về AI: Huấn Luyện Model Chỉ Chiếm 2% Công Việc ML
Sự Thật Ít Ai Nói Về AI: Huấn Luyện Model Chỉ Chiếm 2% Công Việc ML

Nhiều người nghĩ làm AI chủ yếu là train model trên GPU. Thực tế huấn luyện chỉ ~2% — phần lớn thời gian nằm ở ontology, data cleaning và evaluation ML.

machine learningml ontologydata cleaningml evaluation
·4 min read
Tự Host Open-Source LLM Cho AI Agent: Trade-off Chi Phí, Kiểm Soát và Độ Tin Cậy
Tự Host Open-Source LLM Cho AI Agent: Trade-off Chi Phí, Kiểm Soát và Độ Tin Cậy

GLM-5.2, DeepSeek-V4, MiniMax-M3 giờ cạnh tranh agentic. Khi nào tự host tiết kiệm hơn API và những chi phí ẩn hay bị bỏ qua khi xây production agent.

ai agentopen source llmself host llmllm optimization
·7 min read
AWS DevOps Agent: Review Code Sinh Bởi AI Trước Release – Trade-off Thực Tế

AWS DevOps Agent thêm review và test tự động cho code sinh bởi AI. Phân tích giá trị và trade-off áp dụng vào pipeline release thực tế.

ai code reviewdevops agentawsai-agent
·8 min read
Xây AI Agent Production Đáng Tin Cậy: Grounding, Prompt-vs-Code Và Eval Loop Thực Tế
Xây AI Agent Production Đáng Tin Cậy: Grounding, Prompt-vs-Code Và Eval Loop Thực Tế

Agent hay fail production. Phân tích grounding + prompt-vs-code và eval từ signal thực để ship agent giá trị thay vì nợ kỹ thuật.

ai agentproduction aiagent evaluationgrounded agents
·8 min read
Background Agent Production: Governance Và Trade-off Cost, Risk Thực Tế Bạn Phải Biết
Background Agent Production: Governance Và Trade-off Cost, Risk Thực Tế Bạn Phải Biết

OpenAI mua Ona cho governed background agents, Gemini Spark ra mắt 24/7 proactive agent. Phân tích governance layer, persistence, steerability và khi nào autonomous agents tạo ROI thay vì chi phí bất ngờ.

ai agent governancebackground agentpersistent agentgemini spark
·7 min read
DiffusionGemma Local + Hybrid Routing: Cắt 3-5× Chi Phí Cho 70% Subtask Agent Thực Tế

DiffusionGemma 700-1000 t/s local thay đổi kinh tế cho agent. Phân tích hybrid routing cắt 3-5× cost và latency cho developer production.

diffusiongemmalocal llmhybrid llm routingai agent
·8 min read
Claude Fable 5: Cost-Aware Routing Cho Agent Production — Trade-off Thực Tế Bạn Phải Biết
Claude Fable 5: Cost-Aware Routing Cho Agent Production — Trade-off Thực Tế Bạn Phải Biết

Claude Fable 5 mạnh agent dài hơi nhưng giá gấp đôi và burn nhanh. Phân tích routing + cache + cap để đạt ROI thực thay vì chi phí bất ngờ.

claude fable 5ai agentllm cost optimizationmodel routing
·7 min read
Xây Portfolio Tự Cập Nhật Hàng Ngày bằng Grok × Claude
Xây Portfolio Tự Cập Nhật Hàng Ngày bằng Grok × Claude

Pipeline tự động hóa toàn bộ content lifecycle: Grok research real-time → Claude validate schema + guardrail bảo mật → Astro render data-driven blocks → deploy + LinkedIn auto-post. Không phải demo — đây là evidence của hệ thống production với capability routing, security boundary, lazy archive, và mechanism GIF không dùng WebGL.

Agentic PipelineGrokClaudeAstro
·8 min read
Agentic Governance Production: Step Functions + AgentCore Thực Tế
Agentic Governance Production: Step Functions + AgentCore Thực Tế

Governance trong agentic workflow không phải layer thêm sau — đó là state machine bạn phải thiết kế từ ngày 1. Step Functions + AgentCore: retry semantics, Wait for Task Token cho human-in-the-loop không tốn Lambda idle, idempotency contract per step, và Saga compensation cho production-grade agentic pipeline.

AWSStep FunctionsAgentCoreAI Engineering
·10 min read