NemoClaw Knowledge Wiki

Tag: colibri

5 items with this tag.

  • Jul 30, 2026

    model-quantization

    • concept
    • quantization
    • model-compression
    • llm-efficiency
    • bitnet
    • turboquant
    • on-device-deployment
    • moe
    • ram-inference
    • colibri
    • 744b
    • hermes-agent
    • local-ai
    • comfyui
    • int8
    • vram-optimization
    • tts
    • cpu-inference
    • inflect-micro
  • Jul 30, 2026

    sufficient-parameters

    • model-benchmarking
    • small-language-models
    • parameter-efficiency
    • 4gb-models
    • slm-evaluation
    • on-device-ai
    • cognitive-core
    • multimodal-slm
    • edge-ai
    • function-calling
    • moe
    • colibri
    • qwen
    • fablevibes
    • local-llm
    • tts
    • cpu-inference
    • inflect-micro
  • Jul 22, 2026

    inference-optimization

    • inference-speed
    • kv-cache-compression
    • llm-efficiency
    • model-quantization
    • rotorquant
    • context-window
    • tensor-compression
    • gpu-throughput
    • deepseek
    • speculative-decoding
    • paged-attention
    • vram-optimization
    • reasoning-efficiency
    • fine-tuning
    • edge-ai
    • function-calling
    • small-language-models
    • persistent-memory
    • local-ai
    • librarian-system
    • hardware-trade-offs
    • revenue-strategy
    • moe-architecture
    • colibri
    • glom-5.2
    • hermes-agent
    • self-improving-agents
  • Jul 22, 2026

    qwen-36-35b-a3b

    • ai-model
    • qwen
    • moe
    • llm
    • 35b-parameters
    • mixture-of-experts
    • sparse-activation
    • sparse-moe
    • 35b-model
    • edge-deployment
    • low-vram-inference
    • qwen-3.6
    • gemini-3.5-flash
    • google-gemini
    • claude-opus
    • anthropic
    • evaluation-awareness
    • reliability
    • tts
    • miso-tts
    • nvidia-nemotron
    • colibri
    • 744b-model
    • consumer-hardware
    • glom-5.2
  • Jul 14, 2026

    Colibri: Local GLM-5.2 (744B) RAM Inference with MoE, No GPU

    • colibri

Created with Quartz v4.5.2 © 2026

  • GitHub
  • Discord Community