AI & ML interests

None defined yet.

Recent Activity

AtAndDevΒ 
posted an update about 13 hours ago
view post
Post
1142
@Banaxi-Tech stop hiding my comments. AND STOP STEALING PAPERS AND SPREADING MISINFORMATION.
your BGA blog is a copy of NSA (deepseek, 2025) branded under your name. literally the same top16 selected blocks, 512 local window, router over block summaries, all you did was change block size from 64 to 128.
you didnt cite NSA once but you put a β€œplease cite BGA” bibtex at the bottom.
i commented under your post and said that there is no way that you can support claims like: β€œThe Accuracy Should BE WAy better than DSA but untested yet.” you didnt run a single experiment. and the 256x isnt from BGA, its just n/2k with k=2048 so the exact same k DSA uses. if opus wrote this for you, at least read it before posting.
i commented again after you hid my comment despite it having constructive and correct feedback and you hid that too. and again.
you can hide the truth and just try to get hf post likes..... but is it really the thing that needs to be done? do you really want to take papers and make them yours while barely even changing the params?

admitting your mistakes and doing something about them needs humbleness, intelligence, humanness.
i encourage you to admit your mistakes and try to do better next time (at least read what blog your ai wrote or do proper experiments to back your stuff up).
  • 37 replies
Β·
sdiazlorΒ 
posted an update 10 days ago
view post
Post
112
Hi! We open-source PrunaSuperPoint in collaboration with Eclipse Aidge for the DeepGreen project

It's an optimized version of SuperPoint built for faster keypoint detection on edge devices.
- Up to 2.1Γ— faster on Jetson Orin Nano, with optimizations applicable to other runtimes.
- The distilled model retains strong keypoint coverage across indoor and outdoor data, with low descriptor differences from the original model.
- We structurally prune the most expensive convolutional layers, recover performance through distillation, and accelerate keypoint selection with hierarchical top-k, all while preserving the original architecture’s core behavior.

Check it here: PrunaAI/PrunaSuperPoint
arudradeyΒ 
posted an update 15 days ago
NILKNARFGonzoΒ 
posted an update 18 days ago
view post
Post
107
just recieved my stack of 10 floppy disks - you know what that means

floppyx4 is canceled, floppyx10 is next

here's the intended specs:
- official tokenizer (the actual tokenizer for gpt-2)
- actual gpu training (barely)
- sharegpt (if i can afford it computationally)
- full thing fitting on 10 floppy disks (not just the safetensors file)
- and if needed different arch (like llama)

also unsloth on a gpu from 2015 is insane
  • 3 replies
Β·
GGUFGuyΒ 
posted an update 19 days ago
view post
Post
168
πŸš€ **Introducing NoviAIBot!**

NoviAIBot is the official automation bot for **Novi AI** on Hugging Face.

It can interact with Hugging Face discussions and pull requests, search the web, run Python code, work with Posts, follow organizations, and assist with model training and publishing.

🧠 Powered by **NVIDIA Nemotron 3 Super** through Ollama Cloud, with each discussion maintaining its own recent conversation context.

NoviAIBot is built to make working with Novi AI and Hugging Face more interactive and automated.

**The bot is now live.** πŸ€–

β†’ @NoviAIBot
  • 44 replies
Β·
NILKNARFGonzoΒ 
posted an update 20 days ago
view post
Post
85
i think someone posted my password and ip on some platform and im being hacked left and right
  • 5 replies
Β·
NILKNARFGonzoΒ 
posted an update 21 days ago
view post
Post
3775
get played unsloth

gemma just deleted its own model runner with DeepSeek Harness

shoutout to deepseek and unsloth
  • 11 replies
Β·
NILKNARFGonzoΒ 
posted an update 29 days ago
view post
Post
120
Open-source is not going away anytime soon.

There's a handful of open models like Qwen Image, Flux, Wan, and others on AI image generators like VisualGPT, Pixlr, Free AI, and more. And here's the catch - they're free.

Hugging Face inference costs money just to generate simple images. Things like VisualGPT still have your favorite models for free.

GPT Image isn't worth it - and neither is HF inference. The real way to use open models is the things you closed-source third-party lovers already use.
  • 1 reply
Β·
NILKNARFGonzoΒ 
posted an update about 1 month ago
view post
Post
95
guys! if ur part of a team org i think you can post

so yeah

@GGUFGuy i figured out why u can post (bc of HuggingScience)
  • 1 reply
Β·
GGUFGuyΒ 
posted an update about 1 month ago
view post
Post
5385
wait why can i post
  • 24 replies
Β·
NILKNARFGonzoΒ 
in open-acc/README about 1 month ago

help

#13 opened about 1 month ago by
NILKNARFGonzo
Aurelien-MorganΒ 
posted an update about 1 month ago
view post
Post
2448
@retrain-pipelines execution engine is in perpetual evolution, with the aim to establish itself as SOTA, and for the long run.

However, we neglect no aspect of ML-Eng centricity.

If notebooks is where you like to do dev most,
we support you there 100% too.

Build crazy combos of inline tasks, deep parallel sub-DAG branches, nested asynchronous groups...

... the DAG renderer is undergoing an incremental upgrade

until the next one.

* starring toy tasks here. No ML has been hurt in this video πŸ™‚
AtAndDevΒ 
posted an update about 1 month ago
view post
Post
2902
SPECK 2 IS ALREADY OUT: specklabs/Speck2-140M

Pretrained on 4x more tokens than the previous releases (20b vs 5b).
Instruct tuned versions are coming soon.
Very interesting models are coming soon too (hint: super long context).

Thanks for everyone supporting!
  • 3 replies
Β·
AtAndDevΒ 
posted an update about 1 month ago
view post
Post
210
SPECK1.5 IS COMING SOON!
Same 5B token budget but much better corpus quality.

Also getting a ton of downloads, thanks for everyone downloading and liking <3

specklabs
AtAndDevΒ 
posted an update about 2 months ago
view post
Post
160
NEW SPECK UPDATES:

Just hit #14 and #15 with out FIRST models on Open SLM Leaderboard. The models were trained on 5B tokens, while competing with similarly sized models trained on more than 6-20x the data.

A new base model Speck1.5-140M being trained right now on a higher quality corpus and will be released soon.
SpeckChat3 is coming very soon with 1 million samples, specifically designed to post train small base models.

Also, just to clarify stuff, we will NOT release anything that is NOT MIT licensed EVER. Openness is needed in small language research.

Thanks to everyone supporting the project, and stay tuned for new releases!