AI-assisted
deep dive
Review required
Version-sensitive
AI-assisted
Use caution
Overview of multi-node LLM serving architectures comparing vLLM, TensorRT-LLM, and SGLang, with deployment strategies for 70B+ models across GPU clusters.
Read note →
tutorial
Reported tested source
Version-sensitive
Complete walkthrough for installing NVIDIA GPU drivers on Ubuntu, including kernel module removal, secure boot handling, and CUDA toolkit verification.
Read note →
Overview of MiniCPM5-1B, a 1B-parameter on-device language model from 面壁智能, with benchmarks, architecture details, and deployment considerations.
Read note →
tutorial
Reported tested source
Version-sensitive
Offline deployment of MiniCPM5-1B using llama.cpp server in Docker, with Open-WebUI frontend and CPU-only inference configuration.
Read note →
tutorial
Reported tested source
Version-sensitive
Offline deployment of MiniCPM5-1B using Ollama and Open-WebUI in Docker, covering image loading, model import, and CPU-only inference setup.
Read note →
Quick reference for generating an ed25519 SSH key, adding it to the ssh-agent, and configuring it for GitHub authentication on a new machine.
Read note →
Bash script to configure screen blank timeout and power management profiles on Linux laptops, with interactive prompts for quick switching.
Read note →
Graphify turns code, documentation, and papers into a queryable knowledge graph, giving AI coding assistants deeper context across your codebase.
Read note →
Setup and usage notes for Yuxi, an open-source knowledge base and knowledge graph agent platform for building searchable document collections.
Read note →
Overview of OpenCode, an open-source AI coding assistant with intelligent completion, multi-model support, and a CLI-based conversational workflow.
Read note →