Tag
2 articles
This article explains how to fine-tune the Qwen3 language model using Low-Rank Adaptation (LoRA) and NVIDIA NeMo AutoModel in a single-GPU Google Colab environment, focusing on parameter-efficient training techniques and automated workflows.
Perplexity has released pplx-embed, a new collection of multilingual embedding models optimized for large-scale retrieval tasks. The models feature bidirectional attention and diffusion-based training, enhancing their performance in web-scale applications.