Sign in

vb

@reach-vb.hf.co
2.9K followers 103 following 56 posts

GPU Poor @ Hugging Face | F1 fan

PostsRepliesMedia
vb @reach-vb.hf.co · 24/12/2024
You can play directly with the model via this HF space: huggingface.co/spaces/Qwen/...
huggingface.co
QVQ 72B Preview - a Hugging Face Space by Qwen
Discover amazing ML apps made by the community
030
vb @reach-vb.hf.co · 24/12/2024
Model weights here: huggingface.co/Qwen/QVQ-72B...
huggingface.co
Qwen/QVQ-72B-Preview · Hugging Face
We’re on a journey to advance and democratize artificial intelligence through open source and open science.
120
vb @reach-vb.hf.co · 24/12/2024
Qwen released QvQ 72B OpenAI o1 like reasoning model on Hugging Face with Vision capabilities - beating GPT4o, Claude Sonnet 3.5 🔥
4183
vb @reach-vb.hf.co · 06/12/2024
Chat with it live for free here: huggingface.co/chat/models/...
huggingface.co
meta-llama/Llama-3.3-70B-Instruct - HuggingChat
Use meta-llama/Llama-3.3-70B-Instruct with HuggingChat
030
vb @reach-vb.hf.co · 06/12/2024
3.1 70B vs 3.3 70B: Code Generation > HumanEval: 80.5% → 88.4% (+7.9%) > MBPP EvalPlus: 86.0% → 87.6% (+1.6%) Steerability > IFEval: 87.5% → 92.1% (+4.6%) Reasoning & Math > GPQA Diamond (CoT): 48.0% → 50.5% (+2.5%) > MATH (CoT): 68.0% → 77.0% (+9%)
130
vb @reach-vb.hf.co · 06/12/2024
Llama 3.3 70B vs 405B: > GPQA Diamond (CoT): 50.5% vs 49.0% > Math (CoT): 77.0% vs 73.8% > Steerability (IFEval): 92.1% vs 88.6% huggingface.co/meta-llama/L...
huggingface.co
meta-llama/Llama-3.3-70B-Instruct · Hugging Face
We’re on a journey to advance and democratize artificial intelligence through open source and open science.
110
vb @reach-vb.hf.co · 06/12/2024
BOOOOM! Meta released Llama 3.3 70B - 128K context, multilingual, enhanced tool calling, outperforms Llama 3.1 70B and comparable to Llama 405B 🔥 Comparable performance to 405B with 6x LESSER parameters ⚡
1291
vb @reach-vb.hf.co · 03/12/2024
Ofc.. here's the codebase: github.com/huggingface/...
github.com
GitHub - huggingface/parler-tts: Inference and training library for high-quality TTS models.
Inference and training library for high-quality TTS models. - huggingface/parler-tts
020
vb @reach-vb.hf.co · 03/12/2024
And.. here's a space to try out the model too: huggingface.co/spaces/ai4bh...
huggingface.co
Indic Parler-TTS - a Hugging Face Space by ai4bharat
A demo of Indic Parler-TTS
110
vb @reach-vb.hf.co · 03/12/2024
Check out the model checkpoints here: huggingface.co/ai4bharat/in...
huggingface.co
ai4bharat/indic-parler-tts · Hugging Face
We’re on a journey to advance and democratize artificial intelligence through open source and open science.
200
vb @reach-vb.hf.co · 03/12/2024
Introducing Indic-Parler TTS - Trained on 10K hours of data, 938M params, supports 20 Indic languages, emotional synthesis, apache 2.0 licensed! 🔥 w/ fully customisable speech and voice personas! Try it out directly below or use the model weights as you want! 🇮🇳/acc
4353
vb @reach-vb.hf.co · 02/12/2024
try it out today on hf.co/datasets - just click on `SQL Console` followed by `AI Query` 💯
hf.co
Hugging Face – The AI community building the future.
We’re on a journey to advance and democratize artificial intelligence through open source and open science.
260
vb @reach-vb.hf.co · 02/12/2024
you can just do things - ask AI to create your SQL queries and execute them right in your browser! 🔥 let your creativity guide you - powered by qwen 2.5 coder 32b ⚡ available on all 254,746 public datasets on the hub! go check it out today! 🤗
1282
Reposted by vb
Simon Willison @simonwillison.net · 29/11/2024
This demo of structured data extraction running on an LLM that executes entirely in the browser (Chrome only for the moment since it uses WebGPU) is amazing My notes here: simonwillison.net/2024/Nov/29/...
simonwillison.net
Structured Generation w/ SmolLM2 running in browser & WebGPU
Extraordinary demo by Vaibhav Srivastav. Here's Hugging Face's [SmolLM2-1.7B-Instruct](https://huggingface.co/HuggingFaceTB/SmolLM2-1.7B-Instruct) running directly in a web browser (using WebGPU, so r...
518523
vb @reach-vb.hf.co · 28/11/2024
Here's the GitHub repo in case you fancy it: github.com/Vaibhavs10/g...
github.com
GitHub - Vaibhavs10/github-issue-generator-webgpu
Contribute to Vaibhavs10/github-issue-generator-webgpu development by creating an account on GitHub.
080
vb @reach-vb.hf.co · 28/11/2024
To showcase how much you can do with just a 1.7B LLM, you pass free text, define a schema of parsing the text into a GitHub issue (title, description, categories, tags, etc) - Let MLC & XGrammar do the rest! That's it, the code is super readable, try it out today! 🤗 huggingface.co/spaces/reach...
huggingface.co
Github Issue Generator - a Hugging Face Space by reach-vb
Discover amazing ML apps made by the community
1172
vb @reach-vb.hf.co · 28/11/2024
Fuck it! Structured Generation w/ SmolLM2 running in browser & WebGPU 🔥 Powered by MLC Web-LLM & XGrammar ⚡ Define a JSON schema, Input free text, get structured data right in your browser - profit!!
410613
Reposted by vb
Jeremy Howard @howard.fm · 28/11/2024
FYI, here's the entire code to create a dataset of every single bsky message in real time: ``` from atproto import * def f(m): print(m.header, parse_subscribe_repos_message()) FirehoseSubscribeReposClient().start(f) ```
1944262
Reposted by vb
Mark Riedl @markriedl.bsky.social · 26/11/2024
I have converted a portion of my NLP Online Masters course to blog form. This is the progression I present that takes one from recurrent neural network to seq2seq with attention to transformer. mark-riedl.medium.com/transformers...
mark-riedl.medium.com
Transformers: Origins
An unofficial origin story of the transformer neural network architecture.
611615
Reposted by vb
Omar Sanseviero @osanseviero.bsky.social · 27/11/2024
I'm disheartened by how toxic and violent some responses were here. There was a mistake, a quick follow up to mitigate and an apology. I worked with Daniel for years and is one of the persons most preoccupied with ethical implications of AI. Some replies are Reddit-toxic level. We need empathy.
2933237
vb @reach-vb.hf.co · 26/11/2024
> uses 90% sliding window and 10% global attention for efficiency > 2-stage pre-training and 3-phase post-training, including a trapezoid learning rate schedule try it out on hugging face today! 🤗 huggingface.co/collections/...
huggingface.co
Hymba - a nvidia Collection
A series of Hybrid Small Language Models.
040
vb @reach-vb.hf.co · 26/11/2024
yo! nvidia finally released the weights for Hymba-1.5B - outperforms Qwen, and SmolLM2 w/ 6-12x less training trained ONLY on 1.5T tokens > massive reductions in KV cache size and improved throughput > combines Mamba and Attention in a hybrid parallel architecture with a 5:1 ratio and meta-tokens
1292
Reposted by vb
Andi @andimara.bsky.social · 26/11/2024
Let's go! We are releasing SmolVLM, a smol 2B VLM built for on-device inference that outperforms all models at similar GPU RAM usage and tokens throughputs. SmolVLM can be fine-tuned on a Google collab and be run on a laptop! Or process millions of documents with a consumer GPU!
410422
vb @reach-vb.hf.co · 25/11/2024
You can run inference via llama.cpp too: huggingface.co/OuteAI/OuteT...
huggingface.co
OuteAI/OuteTTS-0.2-500M-GGUF · Hugging Face
We’re on a journey to advance and democratize artificial intelligence through open source and open science.
030
vb @reach-vb.hf.co · 25/11/2024
Model weights on the hub, you can even run this on a Raspberry Pi! Go run, inference now! 🐐 huggingface.co/OuteAI/OuteT...
huggingface.co
OuteAI/OuteTTS-0.2-500M · Hugging Face
We’re on a journey to advance and democratize artificial intelligence through open source and open science.
130
vb @reach-vb.hf.co · 25/11/2024
Smol TTS keeps getting better! Introducing OuteTTS v0.2 - 500M parameters, multilingual with voice cloning! 🔥 > Multilingual - English, Chinese, Korean & Japanese > Cross platform inference w/ llama.cpp > Trained on 5 Billion audio tokens > Qwen 2.5 0.5B LLM backbone > Trained via HF GPU grants
5548
vb @reach-vb.hf.co · 25/11/2024
💯
000
vb @reach-vb.hf.co · 25/11/2024
🐐
000
vb @reach-vb.hf.co · 25/11/2024
It depends on what you define long context; I'm fairly confident up until 64K and moderately till 128K, beyond that - I've personally never tested. Most of my observations are based on chat use-cases.
020
vb @reach-vb.hf.co · 25/11/2024
Yeah! @loubnabnl.hf.co & @eliebak.bsky.social are 🐐
020
vb @reach-vb.hf.co · 25/11/2024
If depends on quant type, in most cases as long as you keep all the vocabs + softmaxes + RoPE in full precision it doesn't lobotomise the model as much Using AWQ/ GPTQ - are also a pretty good substitute for CUDA runners, at 4.5 bpw Source: Hugging Chat + TGI team logs assessments
110
vb @reach-vb.hf.co · 25/11/2024
SmolLM - run, pre-train, fine-tune, evaluate SoTA fully open source LM 🔥 Run with Transformers, MLX, Transformers.js, MLC Web-LLM, Ollama, Candle and more! Apache 2.0 licensed codebase - go explore now!
1362
vb @reach-vb.hf.co · 24/11/2024
FWIW - I love reading your takes and have personally found them insightful on numerous occasions!
040
vb @reach-vb.hf.co · 24/11/2024
A lot more got released like, OpenScholar, smoltalk, Hymba, Open ASR Leaderboard and much more.. Can't wait for the next week! 🤗
030
vb @reach-vb.hf.co · 24/11/2024
Jina AI Jina CLIP v2 - general purpose multilingual and multimodal (text & image) embedding model, 900M params, 512 x 512 resolution, matroyoshka representations (1024 to 64) Apple AIM v2 & CoreML MobileCLIP - large scale vision encoders outperform CLIP and SigLIP CoreML optimised MobileCLIP models
130
vb @reach-vb.hf.co · 24/11/2024
Llava o1 - vlm capable of spontaneous, systematic reasoning, similar to GPT-o1, 11B model outperforms gemini-1.5-pro, gpt-4o-mini, and llama-3.2-90B-vision Black Forest Labs Flux.1 tools - four new state of the art model checkpoints & 2 adapters for fill, depth, canny & redux, open weights
120
vb @reach-vb.hf.co · 24/11/2024
All the model checkpoints with URLs here: huggingface.co/posts/reach-...
huggingface.co
@reach-vb on Hugging Face: "Massive week for Open AI/ ML: Mistral Pixtral & Instruct Large - ~123B, 128K…"
We’re on a journey to advance and democratize artificial intelligence through open source and open science.
120
vb @reach-vb.hf.co · 24/11/2024
Massive week for Open Source AI/ ML Mistral Pixtral & Instruct Large - ~123B, 128K context, multilingual, json + function calling & open weights Allen AI Tülu 70B & 8B - competive with claude 3.5 haiku, beats all major open models like llama 3.1 70B, qwen 2.5 and nemotron
3559
vb @reach-vb.hf.co · 24/11/2024
Don’t know about anons, but I’m quite obsessed w/ Ai2 🤗
040
vb @reach-vb.hf.co · 23/11/2024
Very handy! Going to give it a shot tomorrow lol
010
vb @reach-vb.hf.co · 23/11/2024
I mostly thought that one would upsample the prompt w/ a LLM before passing too LTV - hence the level of detail. Probably also the curse of going from a low dimensional (text) representation to high dimensional too.
050
vb @reach-vb.hf.co · 23/11/2024
Hugging Face model repo: huggingface.co/apple/coreml...
huggingface.co
apple/coreml-mobileclip · Hugging Face
We’re on a journey to advance and democratize artificial intelligence through open source and open science.
030
vb @reach-vb.hf.co · 23/11/2024
Apple released blazingly fast CoreML models AND an iOS app to run them on iPhone! ⚡ > S0 matches OpenAI's ViT-B/16 in zero-shot performance but is 4.8x faster and 2.8x smaller > S2 outperforms SigLIP's ViT-B/16 in zero-shot accuracy, being 2.3x faster, 2.1x smaller, and trained with 3x fewer data
2434
vb @reach-vb.hf.co · 23/11/2024
Congratulations on the brilliant release! ❤️ Love that all the models are on Hugging Face and ready to run inference from: huggingface.co/collections/...
huggingface.co
AIMv2 - a apple Collection
A collection of AIMv2 vision encoders that supports a number of resolutions, native resolution, and a distilled checkpoint.
040
vb @reach-vb.hf.co · 23/11/2024
Check out my new swanky handle! 🦋 - Drop your Hugging Face ID in the comments if you want the same!
4190
vb @reach-vb.hf.co · 22/11/2024
LFG!! XGrammar: a lightning fast, flexible, and portable engine for structured generation! 🔥 > Accurate JSON/grammar generation > 3-10x speedup in latency > 14x faster JSON-schema generation and up to 80x CFG-guided generation GG MLC team is literally the best in the game and slept on! ⚡
2524
vb @reach-vb.hf.co · 20/11/2024
🚨 UPDATE: New Whisper based model competing with Nvidia on Open ASR Leaderboard! 🔥 CrisperWhisper aims to transcribe every spoken word exactly as it is, including fillers, pauses, stutters and false starts Whisper Large V3 fine-tune - beats it by roughly ~1 WER margin ⚡ hf.co/spaces/hf-au...
1183
vb @reach-vb.hf.co · 20/11/2024
No open weights *yet*
030
vb @reach-vb.hf.co · 20/11/2024
You can also try it out at chat.deepseek.com
000
vb @reach-vb.hf.co · 20/11/2024
I really hope they release a technical report or the model weights soon! 🤞
110