# LongCat 2.0: Self-Hosting Hardware, Scenarios and Commercial License

> Long-horizon coding, search, repository edits and tool-driven agents on GPU or NPU clusters.

- Verified: 2026-09-12
- Released: 2026-07-02
- Canonical: https://chinaapi.ai/open-model-deployment/longcat-2.0/
- Back to directory: https://chinaapi.ai/open-model-deployment/
- Machine-readable dataset: https://chinaapi.ai/data/open-model-deployment.json

## Model facts

- Vendor: Meituan
- Parameters: 1.6T total; approximately 48B active
- Native precision: Vendor checkpoint; exact serving precision varies by recipe
- Context: 1M tokens
- Category: General and agentic foundation models
- Modalities: text → text
- Serving paths: SGLang, SGLang-FluentLLM
- Official weights: https://huggingface.co/meituan-longcat/LongCat-2.0?utm_source=chinaapi&utm_medium=research&utm_campaign=open-model-deployment
- Model documentation: https://github.com/meituan-longcat/LongCat-2.0?utm_source=chinaapi&utm_medium=research&utm_campaign=open-model-deployment

## Deployment evidence

- **Official launch / minimum — evidence C:** GPU and NPU paths are documented, but the vendor does not publish an absolute minimum hardware shape.
- **200-person team — evidence E:** A 1.6T model needs a measured cluster design; employee count alone is not a capacity input.
- **Commercial API — evidence C:** Official GPU and NPU serving paths exist; topology, redundancy and throughput remain operator-specific. Source: https://github.com/meituan-longcat/LongCat-2.0?utm_source=chinaapi&utm_medium=research&utm_campaign=open-model-deployment

## Scenario evidence

- **Repository-scale coding — Vendor-stated:** Official materials emphasize repository edits and integrations with coding-agent harnesses. Source: https://github.com/meituan-longcat/LongCat-2.0?utm_source=chinaapi&utm_medium=research&utm_campaign=open-model-deployment
- **Search and general agents — Public evidence:** The vendor publishes coding, BrowseComp, RWSearch and agent evaluations. Source: https://github.com/meituan-longcat/LongCat-2.0?utm_source=chinaapi&utm_medium=research&utm_campaign=open-model-deployment
- **NPU-based sovereign deployment — Vendor-stated:** An official SGLang-FluentLLM NPU serving path is linked. Source: https://github.com/meituan-longcat/LongCat-2.0?utm_source=chinaapi&utm_medium=research&utm_campaign=open-model-deployment

## Commercial-use check

- License: MIT — https://github.com/meituan-longcat/LongCat-2.0/blob/main/LICENSE?utm_source=chinaapi&utm_medium=research&utm_campaign=open-model-deployment
- MaaS / hosted service: No model-specific MaaS restriction identified in the MIT license.
- Attribution: Retain the copyright and permission notice.

## Avoid or validate first

- Workstation deployment
- Quoting in-house benchmark results as ChinaAPI reproduction

## Known limitations and open questions

- No official minimum GPU count
- Most published evaluation values are vendor-measured

## Evidence boundary

Loading weights, completing a first forward pass and meeting a production latency/SLA are separate thresholds. The 200-person and commercial tiers require workload-specific measurement. License summaries are product research, not legal advice.

Generated by GENERATED BY scripts/gen_open_model_deployment.py.
