PulseAugur
EN
LIVE 15:05:40

User seeks help with custom micro-llama model integration issues

A user is encountering issues when trying to use a custom-trained micro-llama model in applications like Unsloth Studio. Despite their own implementation successfully generating stories based on their dataset, other applications yield poorer results. The user suspects a configuration or bug in how these applications handle the model's training data, particularly concerning instruction masking and response formatting. They are seeking advice on troubleshooting steps and potential solutions. AI

RANK_REASON User-generated content on a niche subreddit seeking help with a custom model, not a significant industry event.

Read on r/LocalLLaMA →

AI-generated summary · Google Gemini · from 1 sources. How we write summaries →

User seeks help with custom micro-llama model integration issues

COVERAGE [1]

  1. r/LocalLLaMA TIER_1 English(EN) · /u/Darlanio ·

    What am I missing?

    <!-- SC_OFF --><div class="md"><p>Training my own micro-llama-model on a dataset I have published with my own program, I somehow fail to get it working in Unsloth Studio and other applications.</p> <p>The Dataset is here:<br /> <a href="https://huggingface.co/datasets/Darlanio/Sh…