NemoClaw Knowledge Wiki

Tag: Qwen

9 items with this tag.

  • Jul 22, 2026

    q4-1

    • LLM
    • Benchmark
    • Ternary-Bonsai
    • Qwen
    • Quantization
    • speculative-decoding
    • llm-inference
    • model-efficiency
    • 4-bit
  • Jul 22, 2026

    ternary-bonsai-27b

    • LLM
    • Benchmark
    • Ternary-Bonsai
    • Qwen
    • Performance
    • bonsai-27b
    • qwen-comparison
    • quantization
    • benchmarking
  • Jul 22, 2026

    qwen-35b

    • LLM
    • Qwen
    • Benchmark
    • Model-Comparison
    • qwen-35b
    • large-language-model
    • complex-reasoning
  • Jul 18, 2026

    Local LLM-Powered Privacy-Focused OCR App Development Summary Report

    • LocalLLM
    • Qwen
    • AICoding
  • Jul 14, 2026

    Qwen 3.6 27B Local LLM's TitleForge Performance: Replacing Claude Code

    • Qwen
    • LocalLLM
    • AppleSilicon
  • Jul 13, 2026

    ai-innovation

    • AI
    • LLMs
    • Innovation
    • Government-Tech
    • Brazil
    • Qwen
    • Reasoning
    • ai-innovation
    • llm-optimization
    • swireasoning
  • Jul 12, 2026

    qwen-36-27b

    • LLM
    • Qwen
    • Local-Deployment
    • Performance-Benchmark
    • AI-Model
    • 27B-Parameters
    • Agent-Frameworks
    • llm
    • qwen
    • local-inference
    • code-generation
    • agent-frameworks
    • quantization
    • transformer
    • 27b-parameters
    • reasoning-efficiency
    • fine-tuning
  • Jul 08, 2026

    DeepSeek DSpark: LLM Inference Acceleration via Enhanced Speculative Decoding

    • DeepSeek
    • DSpark
    • Qwen
    • Qwen3
    • vLLM
    • LocalAI
    • LLM
    • SpeculativeDecoding
    • DeepSpec
    • RTX5090
    • AIAgents
    • LocalLLM
    • AIArchitects
    • AIBuilder
  • May 10, 2026

    Achieving Fast 35B MoE AI Model Performance on 6GB VRAM with Llama.cpp

    • LocalAI
    • LLM
    • llamacpp
    • Qwen
    • AIonGPU
    • LowVRAM

Created with Quartz v4.5.2 © 2026

  • GitHub
  • Discord Community