Tileward 35B-A3B
Compressed from Qwen3.6-35B-A3B. Serving requests today.
- Weights
- 68.6 GB → 24.5 GB, 64% smaller
- Active
- ~3B of 35B parameters per token
- Context
- 262,144 tokens
A model that does not fit on the card you have costs a second card, a smaller model, or a cloud bill. Tileward Models fits it.
Works with the clients you already run
28 have a recipe to paste, and anything else that speaks MCP works too.
Fit on the hardware you have, stay the model you chose, and run wherever you run it. The first two are measured against the publisher’s own release.
of weights for Tileward 35B-A3B, from 68.6 GB as published: 2.8× smaller, and 1.46× smaller than the publisher’s own FP8 release.
is the error bar on each score, and the gap between our build and the publisher’s FP8 release sits inside it: level on a multiple-choice test.
places it runs unchanged: our hosted API, your VPC, or air-gapped.
Three are served today and one is no longer served. The rest are compressed and ready to serve.
Compressed from Qwen3.6-35B-A3B. Serving requests today.
Compressed from Qwen3.8-27B. Serving requests today.
Served as its publisher shipped it.
Compressed, not served. Hosted or on your hardware, on request.
Compressed, not served. Hosted or on your hardware, on request.
Compressed, not served. Hosted or on your hardware, on request.
Compressed, not served. Hosted or on your hardware, on request.
Compressed, not served. Hosted or on your hardware, on request.
Compressed, not served. Hosted or on your hardware, on request.
No longer served.
Running one of the unserved ones? We can serve it for you, hosted or on your hardware.
Weights, memory, context window and locking for each served build.
Serving requests today. A mixture-of-experts: 256 experts per layer, 8 of them consulted on any one token. It costs about what a 3B model costs to run, and knows what a 35B model knows..
| Measure | Result |
|---|---|
| Base | Qwen3.6-35B-A3B |
| Parameters | 35B total, ~3B active |
| Size on disk | Weights on disk, measured rather than estimated, on the catalog’s MiB/1000 basis. |
| Smaller by | 64%, a 2.8× ratio |
| Context window | 262,144 tokens |
| Per-tile locking | Available |
Measured against the publisher’s FP8 build: the eval · the write-up.
Serving requests today, with text and image input.
| Measure | Result |
|---|---|
| Base | Qwen3.8-27B |
| API id | Tileward-Qwen3.8-27b |
| Architecture | Dense |
| Input | Text and images |
| Compressed by Tileward | Yes |
Third-party open weights under Apache 2.0 that we no longer serve.
| Measure | Result |
|---|---|
| Parameters | 117B total, ~5.1B active |
| Weights | As published, unchanged |
| Compressed by Tileward | No |
| On the API | No longer served |
Third-party open weights under Apache 2.0, served as its publisher shipped it.
| Measure | Result |
|---|---|
| Parameters | 20B total, ~3.6B active |
| Weights | As published, unchanged |
| Compressed by Tileward | No |
| Context window | 8,192 tokens |
| Per-tile locking | Not available on this model |
| Where it fits | High-throughput document ingest |
Tileward 35B-A3B, a mixture-of-experts, lands 2.8× smaller than the weights it was compressed from. The dense Tileward-Qwen3.8-27b lands 1.8×, because more of a dense model stays at full precision.
Change the base URL and your existing client works.
The 215 Tileward Governance tiles apply to the self-hosted model exactly as they do to the hosted API, with audit records retained locally.