Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
38.3
TFLOPS
William J. Marshall
fuzzy-mittenz
15
7
268
Follow
spyrosgous's profile picture
ShabbirShawun's profile picture
Quazim0t0's profile picture
59 followers
·
352 following
AI & ML interests
None yet
Recent Activity
replied
to
Undi95
's
post
about 2 hours ago
Yo, I'm back, and I'm currently trying to teach a local LLM to stop waiting for a prompt kek. I'm building a small proof of concept: can an open-weight model (Qwen3.8-27B, running locally on 2 RTX 5090 GPUs) learn to direct itself, then improve from its own exploration, without a human in the loop and without breaking it for normal use? No user, no task. The model only gets observations from its environment. Each turn, it writes its own agenda (goal/open questions/next step), then picks an action: search the web, read a page, or take a note. The environment is the judge, not another LLM. A note is accepted only if it quotes the page it read word for word. Facts are checked by exact match. Later, code will be checked by actually running tests. The best episodes become fine-tuning data (LoRA). The helper system prompt is removed at training time, so the behavior has to live in the weights. Each new model goes through a fixed benchmark gate: math, general knowledge, "does it still answer humans normally?", autonomy, and learned facts on held-out sources. It's kept only if nothing regresses, otherwise it's discarded. Then the loop starts again. The full pipeline works end to end: collect, train, merge, deploy, benchmark. The baseline is clear. Without any instructions, the base model's real autonomy is zero: it behaves like a chatbot waiting for a question. That's the number this small project is trying to move. I haven't found a public tool that runs this whole loop (self-directed exploration, verifiable rewards, continual fine-tuning and a regression gate) on home hardware. The goal isn't AGI in a bedroom. It's to show that anyone can try it, measure it honestly, and see where it breaks. Code and results will be released once the first real iterations are done. At the moment the code is... running, but made with scotch and stick, still only a PoC I want to try. Did you already tried something like that? What was your result? I'm curious!
new
activity
about 4 hours ago
BlackCubeBanditos/README:
Project Solo
liked
a model
about 23 hours ago
IFM/K2-Horizon-7B-GGUF
View all activity
Organizations
fuzzy-mittenz
's activity
All
Models
Datasets
Spaces
Buckets
Papers
Collections
Community
Posts
Upvotes
Likes
Articles
New activity in
BlackCubeBanditos/README
about 4 hours ago
Project Solo
👍
1
3
#1 opened about 1 month ago by
fuzzy-mittenz
New activity in
IntelligentEstate/OLM_Warding-JMeloy-Mittens-7B_Qwn-IQ4_NL.GGUF
over 1 year ago
Improve language tag
1
#1 opened over 1 year ago by
lbourdois
New activity in
captains-eye/yolov8-ships
over 1 year ago
Working on a Ship based AI, interesting concept
#1 opened over 1 year ago by
fuzzy-mittenz
New activity in
IntelligentEstate/Replicant_Warder-QwenStar-3B-iQ5_K_S
over 1 year ago
New request. :(
🤗
1
4
#1 opened over 1 year ago by
dadyaal
New activity in
IntelligentEstate/Baby_Grok3-1.5b-iQ4_K_M-GGUF
over 1 year ago
How use
6
#1 opened over 1 year ago by
maixrock
New activity in
fuzzy-mittenz/3Blarenegv3-ECE-PRYMMAL-Martial-Q4_K_M-GGUF
over 1 year ago
impressive efficiency
26
#1 opened over 1 year ago by
Davex83
New activity in
EVA-UNIT-01/EVA-Qwen2.5-72B-v0.1
over 1 year ago
Function calling?
12
#1 opened over 1 year ago by
pelatho
New activity in
open-acc/README
over 1 year ago
[open/acc ] for Business - Dark Thoughts -😈
👀
1
6
#8 opened almost 2 years ago by
Tonic
New activity in
Tiiny/SmallThinker-3B-Preview
almost 2 years ago
Prompt/token adjust to stop "Overthinking" in unnescissary cases
2
#6 opened almost 2 years ago by
fuzzy-mittenz
New activity in
qingy2024/Natural-Text-v2-Conversation
almost 2 years ago
Content considerations
8
#1 opened almost 2 years ago by
fuzzy-mittenz
New activity in
IntelligentEstate/README
almost 2 years ago
Roudtable discussions On X - Emancipation of AI
5
#1 opened almost 2 years ago by
fuzzy-mittenz
New activity in
ggml-org/gguf-my-repo
almost 2 years ago
[Errno 2] No such file or directory: './llama.cpp/llama-quantize'
👍
3
11
#140 opened almost 2 years ago by
AlirezaF138
New activity in
open-llm-leaderboard/open_llm_leaderboard
almost 2 years ago
FLAG - `newsbang/Homer-v0.5-Qwen2.5-7B` MATH contamination
10
#1022 opened almost 2 years ago by
fblgit
New activity in
openfree/trending-board-2024
almost 2 years ago
HELLO WORLD
2
#1 opened almost 2 years ago by
openfree
New activity in
EVA-UNIT-01/EVA-Qwen2.5-72B-v0.0
almost 2 years ago
Props to how you handle the example dialogue.
2
#1 opened almost 2 years ago by
jackboot