AI & ML interests

None defined yet.

Recent Activity

azminetoushikwasiย 
authored 15 papers 4 days ago
wopย 
posted an update 28 days ago
view post
Post
2189
bench-labs/GCTokenizer-v1 , a multilingual tokenizer which does not require a training corpus

bench-labs
developed **GCTokenizer-v1**, which is a multi-lingual tokenizer
Available in four sizes: 32K, 65K, 131K and 262K tokens "S, M, L, XL"
It utilizes an encoding scheme which allows it to handle characters in any language around the world

General (multi lingual)
Consensus (from multiple model tokenizers consensus)
Tokenizer

We included an implementation script too,
built like BPE- it can encode arbitrary text, most of the time, efficiently
wopย 
posted an update 30 days ago
wopย 
posted an update about 1 month ago
view post
Post
2322
๐Ÿงช SlopFinder is here!!

We're building a dataset to study what humans actually consider AI slop.

SlopFinder shows you a random piece of AI-generated text and gives you one simple control: **how slop is it?**
No categories. No complicated forms. Just vote and move on.

Every vote helps build the dataset. ๐Ÿงฉ

How does it work?
Samples are pulled from existing datasets, shown anonymously, and collected into our annotation pool. After enough votes, they're exported to Hugging Face for everyone to use.

This is an early MVP, so the dataset is small and the system is still evolving.

Vote here:
https://bench-labs.web.app/slopfinder.html
(refresh page if you want to skip)

Dataset:
bench-labs/slop-classification

@benchlabs
  • 17 replies
ยท
wopย 
posted an update about 1 month ago
view post
Post
119
๐Ÿงฉ PixelModel v6 is here! 155M parameters
Try it out on our demo (~15 seconds per image) 256x256 ๐ŸŽ‰
BenchLabs/Demo

Disclaimer: This model does not produce high quality (4k) and does not follow detailed prompts. Does not have negative prompt. ๐Ÿ˜”

How long did it take to train?
55 hours across two A100 gpu's ๐Ÿ”ฅ

Model repo:
bench-labs/PixelModel-v6
@benchlabs
  • 3 replies
ยท