Issues 共 1425
POC for qwen3.5 / qwen3.6 support optimized for AGX Orin
#3492 · alansrobotlab2 · 2026-05-01
[Bug] Multi-turn chat with Qwen3-0.6B truncates output during thinking phase
#3482 · wopelo · 2026-04-14
[Question] Showcase / question: a board-proven offline language runtime on ESP32-C3, and whether this points to a more extreme form of language-runtime compilation
#3459 · Alpha-Guardian · 2026-03-18
[Model Request] Nemotron-3-nano-4b
#3458 · XomaDev · 2026-03-18
[Tracking] Multi-LoRA Serving
#3446 · babusid · 2026-03-06
[Bug] In the MLC_LLM branch of qwen3-vl, when running the qwen3-vl model, a Segfault error occurred.
#3444 · xifengT · 2026-03-04
[Bug] MLC Chat crashes on Pixel 8 when loading prebuilt Llama 3.2 models
#3442 · Grodoe · 2026-03-03
[Bug] Compiling Qwen3-Embedding model fails
#3404 · SuperMasterBlasterLaser · 2025-12-31
[Bug] In Windows nightly cpu build TVM is not found even it it exists. Could not find module 'E:\virtualenv312\mlc_llm_test\Lib\site-packages\tvm\tvm.dll'
#3403 · SuperMasterBlasterLaser · 2025-12-31
[Model Request]
#3400 · Towhid-Fluidtech · 2025-12-24