AI & ML interests

None defined yet.

Recent Activity

jordanplows  updated a Space about 22 hours ago
Watt-Inference/README
jordanplows  updated a model about 23 hours ago
Watt-Inference/w-1
jordanplows  published a model about 23 hours ago
Watt-Inference/w-1
View all activity

Organization Card

E Inference

Making intelligence as accessible as electricity.

About

E Inference compresses large language models after training, reducing their size while preserving up to 90% of the original model's quality. The result is models that are smaller, cheaper, and faster to run without the steep quality loss typical of compression.

Why

Large models are powerful but expensive to run. For intelligence to work like a utility — available everywhere, to everyone — it needs to run efficiently on far less compute than it does today. That's the problem we're solving.

Status

Early stage. This repository will expand to include a compression toolkit, benchmark results, and usage examples.

Contact

founders@e-inference.com

License

Other

datasets 0

None public yet