0.2

ChatGPTNextWeb/NextChat0.2Mar 1, 2024by jafioti

AI Summary

This release introduces new neural network modules, adds support for the Mistral model, and includes various performance optimizations and fixes.

Key Highlights

  • Added Conv1D and Conv2d neural network modules
  • Integrated support for the Mistral model
  • Implemented Q8 Weight Quantization for performance
  • Improved f16 matmul performance and prompt handling
  • Fixed Swish activation function

New Features

  • Conv1D module
  • Conv2d module
  • Mistral model support
  • Q8 Weight Quantization
  • f16 matmul improvements
  • Swish fix

Full Release Notes

## What's Changed
Everything.

And also:
* Simple Pooling implementation by @TheSeamau5 in https://github.com/jafioti/luminal/pull/4
* Conv1D module by @TheSeamau5 in https://github.com/jafioti/luminal/pull/7
* Conv2d by @TheSeamau5 in https://github.com/jafioti/luminal/pull/8
* Add support for Mistral by @TheSeamau5 in https://github.com/jafioti/luminal/pull/9
* Shell script to download mistral that actually works by @TheSeamau5 in https://github.com/jafioti/luminal/pull/10
* Fix swish by @TheSeamau5 in https://github.com/jafioti/luminal/pull/11
* Minor improvement to f16 matmul, Longer prompt and token generation for testing by @TheSeamau5 in https://github.com/jafioti/luminal/pull/12
* Small Improvements to main by @TheSeamau5 in https://github.com/jafioti/luminal/pull/17
* Q8 Weight Quantization by @jafioti in https://github.com/jafioti/luminal/pull/24


**Full Changelog**: https://github.com/jafioti/luminal/compare/0.1...0.2