# ZeroGPU > ZeroGPU is the compute efficiency layer for AI inference. It runs repeatable, high-volume tasks - classification, extraction, PII redaction, moderation, summarization, routing - on specialized small and nano language models across an edge-powered network, faster and cheaper than centralized GPUs, through one OpenAI-compatible API. ## Docs ## Optional - [Website](https://zerogpu.ai) - [Platform](https://platform.zerogpu.ai/) - [llms.txt](https://docs.zerogpu.ai/llms.txt) - [llms-full.txt](https://docs.zerogpu.ai/llms-full.txt)