--- base_model: - mistralai/Mistral-Nemo-Instruct-2407 --- # Join our Discord! https://discord.gg/BeaverAI ## More than 10000 members strong 💪 A hub for users and makers alike! --- ## Drummer is open for new opportunities (I'm a Software Engineer). Contact me through any of these channels: https://linktr.ee/thelocaldrummer ### Thank you to everyone who subscribed through [Patreon](https://www.patreon.com/TheDrummer). Your support helps me chug along in this brave new world. ### FAQ for those out-of-the-loop
🐶 Who is Drummer? Hi! I'm Drummer. I'm a Software Engineer with experience in JavaScript, Golang, Python, and generally engineering the crap out of things. Why I'm in the AI space: - **Exploration:** Everyone is trying to figure out how AI works and what it's capable of. I am too - just not in creating the smartest, safest model at all costs. - **Upskill:** The world is headed towards AI. It is here to stay. This has been my way of brushing up in this new form of computing challenge. - **Value:** I yearn to create value. I feel satisfaction and fulfillment in providing something meaningful for others. - **Fun:** It's just fun using and making models. It's also fun coming up with theories and realizing them in practice (training AI). I started my tuning venture back in mid-2024 when I wanted to improve its literary capabilities. I've come a long way since then and I have branched out and specialized. Foundational models today are optimized for non-creative uses, and I believe there is a place for AI in creativity and entertainment. I am here to take *the road less traveled by*.
❓ What are my models like? **Bottomline:** My models are usually geared towards creativity, usability, and entertainment! While intelligence, correctness, and problem solving are not my priority, they are still one of many qualities I want in my models. The primary goal is to enhance the experience for users looking to use models for creative uses, and other use cases which require no alignment. In an effort to make it clear to myself and to others what I'm aiming for, I've identified certain qualities that my users often want: Creativity - **Writing:** Does it string together words and sentences in a pleasant & effective way? Does it feel like a writer? - **Dynamism:** How good is the AI at being compelling and intriguing in its storytelling? - **Imagination:** Can the AI navigate through a plethora of possibilities? Can it skirt incoherence and rise up to absolute coherence at the end of it? (Dis)alignment - **Attitude:** Does it refuse in both soft or hard ways? Does it lean towards certain corporate/religious/political ethics & beliefs? How does it see the user and itself? - **Morality:** Does it know ethics? Is its language infected with forced positivity? If not, can it still moralize over difficult & dubious themes? - **Formatting:** How stubborn is it with its established formatting? Can it create effective and novel formats to answer the prompt? Intelligence - **Adherence:** Can it follow instructions? Is it sticking to the prompt? Can it understsand you? - **Knowledge:** Does it know about the world in both fictional and non-fictional way? - **Perception:** Can it handle nuance, complexity, and logic? If it doesn't excel in one of these qualities, or if it's overall mediocre for its size, then I would most likely reiterate until I get something right.
💡 Philosophy A person is defined by the language they use. Not whether they speak in English or German, but in how they perceive reality. Just like how we associate a serial killer as a mind that can't map 'murder' to 'evil', an innocent person is a mind that simply can't imagine 'murder'. They get confused when forced to deal with such subjects. AI's use of language speaks volumes about their 'perception' of reality. If a language model has been skewed and limited to a positive perception, then it's ability to imagine is also limited. Finetuning is an opportunity to adjust and broaden the language. Corporations use it to achieve safety and compliance. I'm here to
--- [Drummer](https://huggingface.co/TheDrummer) proudly presents... # Rocinante XL 16B v1 🚀 ![image](https://cdn-uploads.huggingface.co/production/uploads/65f2fd1c25b848bd061b5c2e/S8xCK4xNmqxbBmaaA5AKH.png) #### Rocinante X... # BUT BIGGER! #### (and better) ## What's New? - Updated prose, better writing, fun dialogue, and robust roleplaying! - Upscaled Nemo 12B to improve learning & stability. ## Usage - Mistral v3 Tekken (NOT v7, REMOVE `[SYSTEM_PROMPT]`) - or Metharme (see Discussion for thoughts & praises. Recommended by many.) - \ \ works - No thinking also works ![Screenshot 2026-01-25 at 7.39.16 PM](https://cdn-uploads.huggingface.co/production/uploads/65f2fd1c25b848bd061b5c2e/Ixd1piCOPfJb9uhugYj2-.png) ![Screenshot 2026-01-25 at 7.39.49 PM](https://cdn-uploads.huggingface.co/production/uploads/65f2fd1c25b848bd061b5c2e/pLe13iZQ4no8M8t82U-bS.png) (Mistral v3 Tekken has no whitespace and no `[SYSTEM_PROMPT]`) ## Description > After using stupid shit like 26B gemma4 finetunes for weeks. going back to roc xl is nice > I am using it currently. With 12b models like cydonia i had problems with the AI not understanding some nuances, even with thinking. This model with thinking pushes way above it's weight in my opinion, feels more like a 24b to 32b model. At least judging from my current brief experience > wow, it writes really well! It's really close to a Cydonia. And, I mean v4.3+ Cydonia, not before. If I was sacrilegious, I'd say close to v4zr. I'll continue using it, it has really a huge potential. But yes, it punches really hard for its size. When I have some time, I'll try to set up reasoning \, I can put some Q4 all into VRAM with a bit of space left. Really curious to see how it goes! > The model prose is way better than 12B - the former often writes in a specific way, and 16B is far more flexible. > I'd like to echo what other people have been saying about Rocinante XL 16B v1's prose, it's actually really solid and often beats out the last Magidonia/Skyfall versions I tried in that regard. > So far it's goated for me. Using meth > it really is like night and day when using metharm vs using teken its not even funny lol > Honestly I'm not using dry or xtc the rest is pretty similar so far I prefer it over most models in the 20B~ range > Yeah iv been bouncing around .8-1.1 and overall it's pretty good! I think logical inconsistencies are the weak point (I still think 12b Rivermind Lux mogs most models even to this day in that front) but the rest is really good it's shocking it's a 16B. Much better than Snowpiercer and I think it holds up well to something like Cydonia. > Been benchmarking Rocinante XL 16B for long-form RP over the last few days and honestly I’m pretty impressed. > > What stands out the most isn’t the prose quality (which is already very good), but the character agency and emotional continuity. The model consistently takes initiative, closes scenes when it makes sense, introduces new directions naturally, and generally feels more like it’s roleplaying a character than just reacting to user input. > > One thing I particularly liked: the model occasionally refuses to follow nonsensical narrative pressure from the user. For example, I had a scene where the character was half asleep and I kept trying to continue the conversation. Instead of forcing more dialogue, Rocinante simply advanced the scene to the next morning. That kind of narrative judgment is surprisingly rare. > > Character consistency has also been excellent. In my tests, emotional conflicts evolved naturally over time instead of resetting every few messages. Motivations felt persistent and consequences actually mattered. > > The main weakness I found is context degradation. Up to around 16k context everything feels fantastic. Around 20k I start noticing drift. Past that, repetition, weird callbacks, and occasional nonsense become increasingly common. By 30k+ the quality drop becomes very noticeable. This isn’t unique to Rocinante, but I figured it was worth mentioning. > > Interestingly, I got much better long-term results by using aggressive narrative summaries (~2k tokens of relationship state, character state, emotional state, and open threads) plus a small recent tail, rather than feeding huge amounts of raw chat history. > > Overall I’d easily rank it among the strongest RP-focused models I’ve tested. It genuinely surprised me several times, which is probably the highest compliment I can give a roleplay model. > > Great work. > its the only real improvement the old nemo had in a year ## Links - Original: https://huggingface.co/TheDrummer/Rocinante-XL-16B-v1 - GGUF: https://huggingface.co/TheDrummer/Rocinante-XL-16B-v1-GGUF - iMatrix (recommended): https://huggingface.co/bartowski/TheDrummer_Rocinante-XL-16B-v1-GGUF - EXL3: https://huggingface.co/ArtusDev/TheDrummer_Rocinante-XL-16B-v1-EXL3 `config-v1a`