1Code2LoRA: Hypernetwork-Generated Adapters for Code Language Models under Software Evolutionleanleft@lemmy.mlEnglish · 18 days0 Comments
1ToMoE: Converting Dense Large Language Models to Mixture-of-Experts through Dynamic Structural Pruningleanleft@lemmy.mlEnglish · 18 days0 Comments
1I built two open-source tools to fix AMD GPU local LLM setup (ROCmFix) and benchmark Vulkan vs HIP performance (InferBench)Xanpavle@lemmy.worldEnglish · 21 days0 Comments