July 2, 2026 · AI
Wiring codex to a Local Qwen — ollama, vLLM, and the developer role
Point OpenAI codex CLI at a local Qwen or gemma in five minutes
#codex
#LLM
#vLLM
#ollama
#vLLM
한국어July 2, 2026 · AI
Point OpenAI codex CLI at a local Qwen or gemma in five minutes
July 2, 2026 · AI
Use google-colab-cli to drive free Colab GPUs from your laptop
June 25, 2026 · AI
MTP and diffusion inference on Gemma 4 and Qwen 3.6, fp8 on one H100
May 20, 2026 · AI
Benchmarking Qwen3.5-9B on Apple Silicon across MLX, llama.cpp, Ollama, omlx, and vLLM Metal — single-request throughput, prefill scaling, decode vs input length, and concurrency response