README / README.md
jordanplows's picture
Update README.md
eaa7b88 verified
|
Raw
History Blame Contribute Delete
758 Bytes

E Inference

Making intelligence as accessible as electricity.

About

E Inference compresses large language models after training, reducing their size while preserving up to 90% of the original model's quality. The result is models that are smaller, cheaper, and faster to run without the steep quality loss typical of compression.

Why

Large models are powerful but expensive to run. For intelligence to work like a utility — available everywhere, to everyone — it needs to run efficiently on far less compute than it does today. That's the problem we're solving.

Status

Early stage. This repository will expand to include a compression toolkit, benchmark results, and usage examples.

Contact

Details coming soon.

License

TBD.