Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
36.3
TFLOPS
EsKa
SerialKicked
126
5
98
P(doom)
50%
Follow
GhostGate's profile picture
Coelicidium's profile picture
Nexesenex's profile picture
38 followers
·
74 following
https://github.com/SerialKicked/Lethe-AI-Sharp/
SerialKicked
AI & ML interests
Working on Lethe AI, a modular, object‑oriented C# library that connects local or remote LLM backends to applications. Feature long term memory, backend agnostic code, web search support, and agentic tasks, all rolled into one neat C# lib.
Recent Activity
reacted
to
Undi95
's
post
with 🔥
1 day ago
Yo, I'm back, and I'm currently trying to teach a local LLM to stop waiting for a prompt kek. I'm building a small proof of concept: can an open-weight model (Qwen3.8-27B, running locally on 2 RTX 5090 GPUs) learn to direct itself, then improve from its own exploration, without a human in the loop and without breaking it for normal use? No user, no task. The model only gets observations from its environment. Each turn, it writes its own agenda (goal/open questions/next step), then picks an action: search the web, read a page, or take a note. The environment is the judge, not another LLM. A note is accepted only if it quotes the page it read word for word. Facts are checked by exact match. Later, code will be checked by actually running tests. The best episodes become fine-tuning data (LoRA). The helper system prompt is removed at training time, so the behavior has to live in the weights. Each new model goes through a fixed benchmark gate: math, general knowledge, "does it still answer humans normally?", autonomy, and learned facts on held-out sources. It's kept only if nothing regresses, otherwise it's discarded. Then the loop starts again. The full pipeline works end to end: collect, train, merge, deploy, benchmark. The baseline is clear. Without any instructions, the base model's real autonomy is zero: it behaves like a chatbot waiting for a question. That's the number this small project is trying to move. I haven't found a public tool that runs this whole loop (self-directed exploration, verifiable rewards, continual fine-tuning and a regression gate) on home hardware. The goal isn't AGI in a bedroom. It's to show that anyone can try it, measure it honestly, and see where it breaks. Code and results will be released once the first real iterations are done. At the moment the code is... running, but made with scotch and stick, still only a PoC I want to try. Did you already tried something like that? What was your result? I'm curious!
replied
to
Undi95
's
post
1 day ago
Yo, I'm back, and I'm currently trying to teach a local LLM to stop waiting for a prompt kek. I'm building a small proof of concept: can an open-weight model (Qwen3.8-27B, running locally on 2 RTX 5090 GPUs) learn to direct itself, then improve from its own exploration, without a human in the loop and without breaking it for normal use? No user, no task. The model only gets observations from its environment. Each turn, it writes its own agenda (goal/open questions/next step), then picks an action: search the web, read a page, or take a note. The environment is the judge, not another LLM. A note is accepted only if it quotes the page it read word for word. Facts are checked by exact match. Later, code will be checked by actually running tests. The best episodes become fine-tuning data (LoRA). The helper system prompt is removed at training time, so the behavior has to live in the weights. Each new model goes through a fixed benchmark gate: math, general knowledge, "does it still answer humans normally?", autonomy, and learned facts on held-out sources. It's kept only if nothing regresses, otherwise it's discarded. Then the loop starts again. The full pipeline works end to end: collect, train, merge, deploy, benchmark. The baseline is clear. Without any instructions, the base model's real autonomy is zero: it behaves like a chatbot waiting for a question. That's the number this small project is trying to move. I haven't found a public tool that runs this whole loop (self-directed exploration, verifiable rewards, continual fine-tuning and a regression gate) on home hardware. The goal isn't AGI in a bedroom. It's to show that anyone can try it, measure it honestly, and see where it breaks. Code and results will be released once the first real iterations are done. At the moment the code is... running, but made with scotch and stick, still only a PoC I want to try. Did you already tried something like that? What was your result? I'm curious!
liked
a model
1 day ago
TheDrummer/Artemis-31B-v1.2
View all activity
Organizations
SerialKicked
's activity
All
Models
Datasets
Spaces
Buckets
Papers
Collections
Community
Posts
Upvotes
Likes
Articles
liked
a model
1 day ago
TheDrummer/Artemis-31B-v1.2
31B
•
Updated
10 days ago
•
1.49k
•
40
liked
a model
24 days ago
zerofata/G4-MeroMero-v2-31B-GGUF
31B
•
Updated
Aug 3
•
10.6k
•
41
liked
a model
about 2 months ago
Qwen/Qwen3.8-27B
Image-Text-to-Text
•
28B
•
Updated
Aug 14
•
6.76M
•
•
17.2k
liked
a model
2 months ago
moonshotai/Kimi-K3
Image-Text-to-Text
•
2.8T
•
Updated
Sep 2
•
1.25M
•
•
11.6k
liked
a model
3 months ago
Gryphe/Gemma-4-31B-StyleTune
Text Generation
•
33B
•
Updated
6 days ago
•
709
•
•
95
liked
5 models
6 months ago
TheDrummer/Artemis-31B-v1-GGUF
31B
•
Updated
Aug 6
•
1.01k
•
27
zerofata/G4-MeroMero-26B-A4B
26B
•
Updated
May 2
•
467
•
126
BeaverAI/Artemis-31B-v1f-GGUF
31B
•
Updated
Apr 17
•
228
•
2
BeaverAI/Artemis-31B-v1b-GGUF
31B
•
Updated
Apr 9
•
419
•
14
google/gemma-4-31B-it
Image-Text-to-Text
•
31B
•
Updated
Jul 20
•
9.59M
•
•
4.05k
liked
4 models
7 months ago
zerofata/Q3.5-BlueStar-v2-27B
27B
•
Updated
Mar 20
•
43
•
45
ACE-Step/Ace-Step1.5
Text-to-Audio
•
Updated
Feb 3
•
67.6k
•
889
zerofata/Q3.5-BlueStar-27B-gguf
27B
•
Updated
Mar 3
•
200
•
15
Qwen/Qwen3.5-27B
Image-Text-to-Text
•
28B
•
Updated
Apr 24
•
1.92M
•
•
1.06k
liked
3 models
8 months ago
aixonlab/Eurydice-24b-v3.5
Text Generation
•
24B
•
Updated
May 23, 2025
•
29
•
•
11
zerofata/MS3.2-PaintedFantasy-v4-24B
24B
•
Updated
Feb 7
•
174
•
21
allenai/Bolmo-7B
Text Generation
•
8B
•
Updated
about 16 hours ago
•
378
•
59
liked
2 models
9 months ago
allenai/Olmo-3.1-32B-Think
Text Generation
•
32B
•
Updated
Jan 5
•
19.1k
•
115
TheDrummer/Cydonia-24B-v4.3
415k
•
Updated
Dec 17, 2025
•
2.43k
•
171
liked
a model
10 months ago
NousResearch/Hermes-4.3-36B
Text Generation
•
36B
•
Updated
Dec 6, 2025
•
1.57k
•
306
Load more