SCI AI AL Alain Airom (Ayrom) Run Big LLMs on Small GPUs: A Hands-On Guide to 4-bit Quantization and QLoRA Save the planet and adapt the LLM to your use-case!
AI AL Alain Airom (Ayrom) Shrinking Giants: A Word on Floating-Point Precision in LLM Domain for Faster, Cheaper Models Ever wondered how floating-point decision can have an impact on LLM’s output?
AI TE Tech-Practice What makes the Nvidia 5090/5080 unique? Nvidia just released their new 5000 GPUs. How are they different from previous generation cards?