amd/Phi-4-mini-instruct-w4a16-llmcompressor Text Generation • 4B • Updated 2 days ago • 167 • 1
amd/Phi-4-mini-instruct-w4a16-llmcompressor Text Generation • 4B • Updated 2 days ago • 167 • 1
amd/Instella-MoE-16B-A3B-SFT-w8a8-llmcompressor Text Generation • 16B • Updated 2 days ago • 312
amd/Instella-MoE-16B-A3B-SFT-w8a8-llmcompressor Text Generation • 16B • Updated 2 days ago • 312
Evaluating Embedding APIs for Information Retrieval Paper • 2305.06300 • Published May 10, 2023 • 1
NoMIRACL: Knowing When You Don't Know for Robust Multilingual Retrieval-Augmented Generation Paper • 2312.11361 • Published Dec 18, 2023 • 1
On the importance of Data Scale in Pretraining Arabic Language Models Paper • 2401.07760 • Published Jan 15, 2024 • 1
Beyond the Limits: A Survey of Techniques to Extend the Context Length in Large Language Models Paper • 2402.02244 • Published Feb 3, 2024 • 1
QDyLoRA: Quantized Dynamic Low-Rank Adaptation for Efficient Large Language Model Tuning Paper • 2402.10462 • Published Feb 16, 2024
When Chosen Wisely, More Data Is What You Need: A Universal Sample-Efficient Strategy For Data Augmentation Paper • 2203.09391 • Published Mar 17, 2022 • 1