File size: 758 Bytes
eaa7b88
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
1
2
3
4
5
6
7
8
9
10
11
12
13
14
15
16
17
18
19
20
21
22
23
24
25

# E Inference

Making intelligence as accessible as electricity.

## About

E Inference compresses large language models after training, reducing their size while preserving up to 90% of the original model's quality. The result is models that are smaller, cheaper, and faster to run without the steep quality loss typical of compression.

## Why

Large models are powerful but expensive to run. For intelligence to work like a utility — available everywhere, to everyone — it needs to run efficiently on far less compute than it does today. That's the problem we're solving.

## Status

Early stage. This repository will expand to include a compression toolkit, benchmark results, and usage examples.

## Contact

Details coming soon.

## License

TBD.