16 1 109

Simon Salmon

BigSalmon

AI & ML interests

None yet

Recent Activity

liked a model 14 days ago

deepseek-ai/DeepSeek-R1-Distill-Qwen-1.5B

liked a model 28 days ago

tiiuae/Falcon3-3B-Base-1.58bit

liked a model 28 days ago

tiiuae/Falcon3-1B-Base

View all activity

Organizations

BigSalmon's activity

liked a model 14 days ago

deepseek-ai/DeepSeek-R1-Distill-Qwen-1.5B

Text Generation • Updated 7 days ago • 526k • • 710

liked 2 models 28 days ago

tiiuae/Falcon3-3B-Base-1.58bit

Text Generation • Updated about 1 month ago • 167 • 2

tiiuae/Falcon3-1B-Base

Text Generation • Updated Dec 17, 2024 • 10.1k • 20

liked a model 2 months ago

mlx-community/Llama-3.3-70B-Instruct-3bit

Text Generation • Updated Dec 6, 2024 • 1.55k • 6

reacted to merve's post with 🔥 2 months ago

Post

2919

Last week we were blessed with open-source models! A recap 💝
merve/nov-29-releases-674ccc255a57baf97b1e2d31

🖼️ Multimodal
> At Hugging Face we released SmolVLM, a performant and efficient smol vision language model 💗
> Show Lab released ShowUI-2B: new vision-language-action model to build GUI/web automation agents 🤖
> Rhymes AI has released the base model of Aria: Aria-Base-64K and Aria-Base-8K with their respective context length
> ViDoRe team released ColSmolVLM: A new ColPali-like retrieval model based on SmolVLM
> Dataset: Llava-CoT-o1-Instruct: new dataset labelled using Llava-CoT multimodal reasoning model📖
> Dataset: LLaVA-CoT-100k dataset used to train Llava-CoT released by creators of Llava-CoT 📕

💬 LLMs
> Qwen team released QwQ-32B-Preview, state-of-the-art open-source reasoning model, broke the internet 🔥
> AliBaba has released Marco-o1, a new open-source reasoning model 💥
> NVIDIA released Hymba 1.5B Base and Instruct, the new state-of-the-art SLMs with hybrid architecture (Mamba + transformer)

⏯️ Image/Video Generation
> Qwen2VL-Flux: new image generation model based on Qwen2VL image encoder, T5 and Flux for generation
> Lightricks released LTX-Video, a new DiT-based video generation model that can generate 24 FPS videos at 768x512 res ⏯️
> Dataset: Image Preferences is a new image generation preference dataset made with DIBT community effort of Argilla 🏷️

Audio
> OuteAI released OuteTTS-0.2-500M new multilingual text-to-speech model based on Qwen-2.5-0.5B trained on 5B audio prompt tokens