
Exla
An SDK to run transformer models anywhere
San Francisco , US
Artificial Intelligence
Edge Computing Semiconductors
Computer Vision
Exla aggressively quantizes AI models to minimize memory usage and maximize inference speed. Whether you're deploying LLMs, VLMs, VLAs, or custom models, Exla reduces memory footprint by up to 80% and accelerates inference by 3–20x - all with just a few lines of code.
https://cal.com/exla-ai/schedule
- Projects
- Team
- Jobs
- News
No projects yet.
This company hasn't published any projects. Be the first to launch one!
Create Your First ProjectFounded2025
Team Size2
LocationUS
Websiteexla.ai/
🏳️
Claim This ListingIs this your company?
Claim this listing to manage your profile, add updates, and connect with your audience.