AI & ML interests

None defined yet.

Recent Activity

jordanplows  updated a model about 10 hours ago
Watt-Inference/w-1
jordanplows  updated a Space about 10 hours ago
Watt-Inference/README
jordanplows  published a model 4 days ago
Watt-Inference/w-1
View all activity

Organization Card

Watt Inference

Making intelligence as accessible as electricity.

About

Watt Inference compresses large language models after training, reducing their size while preserving up to 90% of the original model's quality. The result is models that are smaller, cheaper, and faster to run without the steep quality loss typical of compression.

Why

Large models are powerful but expensive to run. For intelligence to work like a utility — available everywhere, to everyone — it needs to run efficiently on far less compute than it does today. That's the problem we're solving.

Status

Early stage. This repository will expand to include a compression toolkit, benchmark results, and usage examples.

Contact

founders@wattinference.com

License

Other

datasets 0

None public yet