Comments (2)
It is our company policy not to talk about our future plans until we launch them. We prefer to speak through our work rather than our words. We are working on many exciting projects right now, and we hope you will find them valuable when they launch.
from llm-foundry.
And please make it multilingual
from llm-foundry.
Related Issues (20)
- Fly tokenization with multiple streams HOT 12
- Setting Dropout in MPT Prefix-LM after Exporting to HuggingFace Crashes during Fine-tuning HOT 2
- Evaluation for long_context_tasks failed with a KeyError: 'continuation_indices' HOT 3
- Can you add the pre-training of dbrx? HOT 3
- Installation issue from habana_alpha branch HOT 2
- Fine-tune dbrx-instruct on a single VM with 8 H100s HOT 1
- Is there a way to figure out what dependencies are installed in the docker image? HOT 1
- Opt-3b Pretrain YAML config failing with mosaicml/llm-foundry/2.2.1_cu121_flash2-4aef5de docker HOT 1
- Observing 1/2 the throughput on AMD MI250 HOT 4
- Add State Space Models / Mamba Layer Support
- Possibility of training with hostname instead IP HOT 1
- Train with attention mask HOT 1
- MoE with FSDP HOT 1
- Conversion Sharded -> Monolithic checkpoint HOT 1
- Finetuning does not work on nightly HOT 1
- How to run inference/convert_composer_to_hf.py with MPT-1B model on Habana Gaudi 2, file formats do not match HOT 4
- `ValueError` when following finetuning `mpt-7b-arc-easy--gpu.yaml` example with different default batch size HOT 2
- Issue when installing "pip install -e ".[gpu-flash2]"" HOT 3
- Wrong number of samples for C4? HOT 2
- Composer crashes when attempting to load sharded checkpoint HOT 3
Recommend Projects
-
React
A declarative, efficient, and flexible JavaScript library for building user interfaces.
-
Vue.js
🖖 Vue.js is a progressive, incrementally-adoptable JavaScript framework for building UI on the web.
-
Typescript
TypeScript is a superset of JavaScript that compiles to clean JavaScript output.
-
TensorFlow
An Open Source Machine Learning Framework for Everyone
-
Django
The Web framework for perfectionists with deadlines.
-
Laravel
A PHP framework for web artisans
-
D3
Bring data to life with SVG, Canvas and HTML. 📊📈🎉
-
Recommend Topics
-
javascript
JavaScript (JS) is a lightweight interpreted programming language with first-class functions.
-
web
Some thing interesting about web. New door for the world.
-
server
A server is a program made to process requests and deliver data to clients.
-
Machine learning
Machine learning is a way of modeling and interpreting data that allows a piece of software to respond intelligently.
-
Visualization
Some thing interesting about visualization, use data art
-
Game
Some thing interesting about game, make everyone happy.
Recommend Org
-
Facebook
We are working to build community through open source technology. NB: members must have two-factor auth.
-
Microsoft
Open source projects and samples from Microsoft.
-
Google
Google ❤️ Open Source for everyone.
-
Alibaba
Alibaba Open Source for everyone
-
D3
Data-Driven Documents codes.
-
Tencent
China tencent open source team.
from llm-foundry.