mesh.applied

Task-specialized open models for your local inference needs.

We take the latest open source models, and through a series of post-training, quantization and kernel specialization techniques, we make them run fast on local hardware.

DeepSeek-V4.1-Flash-Ternary
  • - Recommended hardware: 4x RTX6000
  • - 4 users, 276 tok/sec aggregate
  • - for complex coding and research
Qwen3.8-Flash-Next-NVFP4
  • - Recommended hardware: 1x RTX6000
  • - 4 users, 305 tok/sec aggregate
  • - for coding and everyday tasks

with the Goethe framework:

~ git clone https://github.com/meshapplied/goethe.git