0.2
ChatGPTNextWeb/NextChat0.2Mar 1, 2024by jafioti
AI Summary
This release introduces new neural network modules, adds support for the Mistral model, and includes various performance optimizations and fixes.
Key Highlights
- Added Conv1D and Conv2d neural network modules
- Integrated support for the Mistral model
- Implemented Q8 Weight Quantization for performance
- Improved f16 matmul performance and prompt handling
- Fixed Swish activation function
New Features
- Conv1D module
- Conv2d module
- Mistral model support
- Q8 Weight Quantization
- f16 matmul improvements
- Swish fix
Full Release Notes
## What's Changed Everything. And also: * Simple Pooling implementation by @TheSeamau5 in https://github.com/jafioti/luminal/pull/4 * Conv1D module by @TheSeamau5 in https://github.com/jafioti/luminal/pull/7 * Conv2d by @TheSeamau5 in https://github.com/jafioti/luminal/pull/8 * Add support for Mistral by @TheSeamau5 in https://github.com/jafioti/luminal/pull/9 * Shell script to download mistral that actually works by @TheSeamau5 in https://github.com/jafioti/luminal/pull/10 * Fix swish by @TheSeamau5 in https://github.com/jafioti/luminal/pull/11 * Minor improvement to f16 matmul, Longer prompt and token generation for testing by @TheSeamau5 in https://github.com/jafioti/luminal/pull/12 * Small Improvements to main by @TheSeamau5 in https://github.com/jafioti/luminal/pull/17 * Q8 Weight Quantization by @jafioti in https://github.com/jafioti/luminal/pull/24 **Full Changelog**: https://github.com/jafioti/luminal/compare/0.1...0.2