Hugging Face's logo Hugging Face
  • Models
  • Datasets
  • Spaces
  • Buckets new
  • Docs
  • Enterprise
  • Pricing
    • Website
      • Tasks
      • HuggingChat
      • Collections
      • Languages
      • Organizations
    • Community
      • Blog
      • Posts
      • Daily Papers
      • Hardware
      • Learn
      • Discord
      • Forum
      • GitHub
    • Solutions
      • Team & Enterprise
      • Hugging Face PRO
      • Enterprise Support
      • Inference Providers
      • Inference Endpoints
      • Storage Buckets

  • Log In
  • Sign Up
pipenetwork 's Collections
Qwen3.8-2.4T-A95B MLX (REAP)
Muse-Glimmer 30B MLX
MiniMax-H3 MLX
Inkling-Small · MLX
Inkling · MLX
DeepSeek-V4-Flash · MLX
Nemotron TwoTower · MLX
Qwen3.6-35B-A3B — MLX
Ornith-1.0-397B — MLX
GLM-5.2 REAP (expert-pruned)
GLM-5.2 MLX
VISTA MLX
Rio-3.1-Open-30B MLX
Gemma-4-26B-A4B-it MLX
Gemma-4-31B-it MLX
Kimi-K2.7-Code MLX
MiniMax-M3 MLX
Holo-3.1 MLX (computer-use)
Frog (SWE/debugging) MLX
Nemotron-3 MLX (Apple Silicon)

Kimi-K2.7-Code MLX

updated Jun 13

MLX build of Kimi-K2.7-Code. Base is natively 4-bit (int4 experts + bf16 rest); this keeps experts at 4-bit and lifts non-expert layers to 6-bit.

Upvote
-

  • pipenetwork/Kimi-K2.7-Code-MLX-4bit-hiprec

    Text Generation • 1T • Updated Jun 13 • 1.37k • 1

    Note experts@3-bit, attention/router/embeds/shared/dense@6-bit · fits 512GB · smoke-tested

Upvote
-
  • Collection guide
  • Browse collections
Company
TOS Privacy About Careers
Website
Models Datasets Spaces Pricing Docs