TypeAR
Type-safe one-decision-per-token decoding engine for autoregressive LLMs, inspired by Jev.
Type-Safe Decoding for Autoregressive LLMs
Getting started
Use SGLang to configure and serve a compatible autoregressive model on your local GPU server. This example uses Qwen3.8-27B; follow the Qwen3.8-27B SGLang deployment guide to start it with prefix caching enabled.
- Language
- Python 100%
- License
- Apache-2.0
- Latest release
- Release 0.1.2
- Last activity
- Sep 2026
- Added
- 2026-09-17
- Stars at snapshot
- 28
- Status
- Community listing
Source: community catalog. Details and metrics reflect the published community snapshot and are not a performance endorsement. Documentation details extracted from GitHub · the README · the project website.