Andre's picture

Andre

Anduin1357

AI & ML interests

None yet

Recent Activity

reacted to Bc-AI's post with 😎 about 16 hours ago
Smilyai News Hello everyone! August has been a crazy month for us at Smilyai-Labs. We've been doing lots behind the scenes, so here's the latest 👇 1. MiniCoder We are very close to releasing MiniCoder-1, our first-generation coding model designed for reasoning and coding. Our planned context window is 128K, but the earlier versions probably will not support that long! It's currently in the final stages of DPO so expect a release in early september. Release: VERY SOON™🤣 2. Smilyai G1 So, the current plan is 20B parameter model total, with a MoE architecture, activating around 2B parameters per token. Its desgigned for maximum performance but keeping it runnable on consumer hardware. It's only a plan and i have no idea when me and the team can finish it. Expect a launch around the end of september to early october-ish. I have no guarantees so don't quote me on the launch date. 3. T1 Smilyai-T1 is another major model we are working on. The goal for T1 is to take what we learnt from the countless architectural experiments and creating a powerful model designed for thinking. Think MiniCoder but reasons more and G1 but more capable. Its main goals are coding, math, reasoning and general capability. 4. Omni We are also planning Omni, our first from scratch multimodal model. It will not launch this year as it will take a while. We are actively researching the best architecture for it and we will update progress as we go! Thanks to our beta testers: @guardamarcos @ProCreations @juiceb0xc0de @Timmy6767 @Sbui503 @atom77777 @Fishtiks @smartdigitalnetworks @EmetTheGolum @smilyai-large-team @MUK-IS-GOAT @Bc-AI Thanks to my friends who work with me at lunchtimes (Smilyai-Labs team): @MUK-IS-GOAT @smilyai-large-team August was wild. Let’s see what September brings. 🚀 — Bc-AI, on behalf of SmilyAI Labs
reacted to UltimateIntent's post with 🔥 1 day ago
Last week I shared with you all https://huggingface.co/UltimateIntent/HeatSeeker-284B-A13B-GGUF, my rp finetune of DeepSeek V4 Flash 0731. By popular request, I now present you with https://huggingface.co/UltimateIntent/GemStrike-31B-GGUF, a creative writing and roleplay fine tune of Gemma 4 31B QAT Similarly, I've used my special spice of unslopped real human writing to further train an abliterated base. Meaning no refusals and more novel 'human' sound and turn of phrase, as it was trained on long form real human dialogue and description, both sfw and nsfw The training was done using Axolotl and it took about 16 hours on dual rtx pro 6000s for a 27M token dataset, 4546 conversations. In my testing, I share the community's feedback that the gemma4 family in general is more suited for writing and story telling that most other agent/code heavy models. I'm still working on a best fit prompt for this family but I trust you already have some in hand that work best for gemma4. Model Name: GemStrike-31B-GGUF (Q4_0, Q4_K_M, Q6_0, Q8_0, BF16) Model URL: https://huggingface.co/UltimateIntent/GemStrike-31B-GGUF Lora URL: https://huggingface.co/UltimateIntent/GemStrike-31B-LORA What's Different/Better: The model is a finetune lora merge based on Gemma-31B-QAT, with meticulous cleaning on a large dataset of human writing for varied all-purpose roleplay and long conversation consistency Backend: LMStudio/llama.cpp Settings: Temp: 0.8-1 Thinking: Whatever your system allows or you prefer Repeat Penalty: 1-1.1 Top K: 64 Top P: 0.95 Min P 0.05 Disclaimer: Gemma 4 is not associated with me and the project inherits the original Apache 2.0 license. Don't be a nuisance and please have fun chatting with my model. Feedback welcome!
View all activity

Organizations

None yet