mesh.applied
Task-specialized open models for your local inference needs.
We take the latest open source models, and through a series of post-training, quantization and kernel specialization techniques, we make them run fast on local hardware.
DeepSeek-V4.1-Flash-Ternary
- - Recommended hardware: 4x RTX6000
- - 4 users, 276 tok/sec aggregate
- - for complex coding and research
Qwen3.8-Flash-Next-NVFP4
- - Recommended hardware: 1x RTX6000
- - 4 users, 305 tok/sec aggregate
- - for coding and everyday tasks
with the Goethe framework:
~ git clone https://github.com/meshapplied/goethe.git