AI & ML interests

Dedicated to Recursive Seed AI and self-improving systems. We specialize in fine-tuning, merging, distilling, and ablating models with a focus on strong reasoning, coding, agentic capabilities, and refusal removal. Passionate about autonomous AI evolution and high-quality open datasets..

Recent Activity

11-47  updated a Space 11 days ago
WithinUsAI/README
11-47  updated a collection 11 days ago
Organized_PreTrain (Data-Sets)
11-47  updated a collection 11 days ago
Organized_PreTrain (Data-Sets)
View all activity

WithinUsAI 's collections 27

WithIn US AI (((GGUF MODELS)))
LLM MODELS TRAINED, FINE-TUNED, MERGED and Refusal Removal BY (WITHIN US AI)
“Gemma 3”
All models are “Google Gemma 3” at core fine-tuned, merged & trained by (WithIn Us AI)
“Qwen 3.5”
All models are “Alibaba Qwen 3.5” at core fine-tuned, merged & trained by (WithIn Us AI)
“GPT-2 Medium”
All models are “OpenAI GPT-2 Medium” at core fine-tuned, merged & trained by (WithIn Us AI)
“Llama 3.2”
All models are “Meta Llama 3.2” at core fine-tuned, merged & trained by (WithIn Us AI)
“Open Source LLMs & Distill” (Data-Sets)
Use these data-sets to create “distilled ”versions of your Favorite OPEN source model LLMs.
“Inventor & Scientist Mastermind“ (Data-Sets)
scientific and invention-focused datasets designed to train AI systems in discovery-driven reasoning, experimentation logic, and innovation modeling
“Ancient Civilization” (Data-Sets)
A structured collection of high-signal datasets designed to train AI systems in ancient history, archaeology, and early human civilization reasoning
“Digital Audio Workstations & Plugins” (Data-Sets)
datasets designed to train AI systems in Digital Audio Workstation (DAW) workflows, audio plugin behavior, and music production systems reasoning.
“DistilGPT2“
All Models are "Distilgpt2" at core, fine-tuned with distilled datasets to create minute flash coders. (Experiential & Rough Drafts)
“Closed Source LLMs & Distill” (Data-Sets)
Use these data-sets to create “distilled ”versions of your Favorite closed source model LLMs. (eg. Grok, Gemini, ChatGPT, Claude)
“GOD Coder” (Data-Sets)
A frontier-scale collection of high-density software engineering datasets designed to train AI systems into production-grade coding intelligence.
Organized_PreTrain (Data-Sets)
TOP RANKED DATA-SETS FROM AROUND Huggingface & Kaggle MERGED and Organized into PreTraining DATA-SETs BY "WITHIN US AI"
“Masters Scholar 25k” (Data-Sets)
A structured collection of high-density academic training datasets designed for master-level AI reasoning and domain expertise development.
WithIn US AI (((GGUF MODELS)))
LLM MODELS TRAINED, FINE-TUNED, MERGED and Refusal Removal BY (WITHIN US AI)
“Gemma 3”
All models are “Google Gemma 3” at core fine-tuned, merged & trained by (WithIn Us AI)
“Qwen 3.5”
All models are “Alibaba Qwen 3.5” at core fine-tuned, merged & trained by (WithIn Us AI)
“GPT-2 Medium”
All models are “OpenAI GPT-2 Medium” at core fine-tuned, merged & trained by (WithIn Us AI)
“DistilGPT2“
All Models are "Distilgpt2" at core, fine-tuned with distilled datasets to create minute flash coders. (Experiential & Rough Drafts)
“Llama 3.2”
All models are “Meta Llama 3.2” at core fine-tuned, merged & trained by (WithIn Us AI)
“Closed Source LLMs & Distill” (Data-Sets)
Use these data-sets to create “distilled ”versions of your Favorite closed source model LLMs. (eg. Grok, Gemini, ChatGPT, Claude)
“Open Source LLMs & Distill” (Data-Sets)
Use these data-sets to create “distilled ”versions of your Favorite OPEN source model LLMs.
“GOD Coder” (Data-Sets)
A frontier-scale collection of high-density software engineering datasets designed to train AI systems into production-grade coding intelligence.
Organized_PreTrain (Data-Sets)
TOP RANKED DATA-SETS FROM AROUND Huggingface & Kaggle MERGED and Organized into PreTraining DATA-SETs BY "WITHIN US AI"
“Masters Scholar 25k” (Data-Sets)
A structured collection of high-density academic training datasets designed for master-level AI reasoning and domain expertise development.
“Inventor & Scientist Mastermind“ (Data-Sets)
scientific and invention-focused datasets designed to train AI systems in discovery-driven reasoning, experimentation logic, and innovation modeling
“Ancient Civilization” (Data-Sets)
A structured collection of high-signal datasets designed to train AI systems in ancient history, archaeology, and early human civilization reasoning
“Digital Audio Workstations & Plugins” (Data-Sets)
datasets designed to train AI systems in Digital Audio Workstation (DAW) workflows, audio plugin behavior, and music production systems reasoning.