monsoon-nlp

👁 Image
eramax's profile picture 👁 Image
khairi's profile picture 👁 Image
iDrops's profile picture

AI & ML interests

biology and multilingual models

Recent Activity

liked a model 22 days ago

reacted to mmhamdy's post with 🚀 22 days ago

Human brains don't recreate every pixel to understand the world! Most current models in genomics, proteomics, and single-cell transcriptomics rely on generative objectives like masked language modeling or next token prediction. While effective, these architectures waste significant capacity reconstructing raw, noisy sequence details that may not carry functional biological meaning. But a promising, more efficient alternative is emerging: Joint-Embedding Predictive Architecture (JEPA) Originally introduced by Yann LeCun for computer vision, JEPA is a non-generative, self-supervised learning (SSL) framework. Instead of predicting raw inputs, it operates as a world model that predicts abstract semantic embeddings in latent space. Recently, the JEPA framework (and its more efficient LeJEPA variant) has been adapted into the biological sciences to develop performing foundation models and to improve on already existing ones. It's interesting how each adaptation modified and tailored JEPA to suit its specific biological domain, whether by experimenting with different backbones or complementing the objective with other loss terms. For example, JEPA-DNA and ProteinJEPA used JEPA as a continual pre-training framework to enhance existing foundation models without training from scratch, while Cell-JEPA and JEPA-DNA employed a hybrid objective that combines the JEPA loss with a traditional language modeling loss. The article below provides an overview of these implementations, along with others that came out this year. As always, your thoughts and feedback are welcome and highly appreciated! Link to the article is in the first comment 👇

reacted to pankajpandey-dev's post with 🔥 29 days ago

🇮🇳 Qwen3-4B Hindi Instruct v2 — a Hindi LLM that runs on your own machine Most strong Hindi-capable models are either huge or cloud-only. I wanted one that's small enough to run locally but actually follows instructions in Hindi — so I fine-tuned Qwen3-4B on 10K Hindi instruction pairs and shipped it with a full GGUF quant ladder. ✅ Fine-tune (16-bit): huggingface.co/pankajpandey-dev/Qwen3-4B-Hindi-Instruct-v2 ✅ GGUF (Q4/Q5/Q8): huggingface.co/pankajpandey-dev/Qwen3-4B-Hindi-Instruct-v2-GGUF Runs in Ollama, llama.cpp, and LM Studio. The Q4_K_M is just 2.5 GB — fits comfortably on a laptop, CPU or GPU. Part of my Hindi LLM Series — building openly-licensed Indic models for local and edge use. More coming (Gemma next). Feedback welcome 🙏 #Hindi #IndicNLP #GGUF #LocalLLM #Qwen

View all activity

Organizations

👁 BigScience Workshop's profile picture
👁 Spaces-explorers's profile picture
👁 BigCode's profile picture
👁 Blog-explorers's profile picture
👁 Scary Snake's profile picture
👁 Hugging Face Discord Community's profile picture
👁 Hugging Face Context Course's profile picture

👁 Image

liked a model 22 days ago

Updated Feb 24 • 5 • 2

👁 Image

reacted to mmhamdy's post with 🚀 22 days ago

view post

Human brains don't recreate every pixel to understand the world!

Most current models in genomics, proteomics, and single-cell transcriptomics rely on generative objectives like masked language modeling or next token prediction. While effective, these architectures waste significant capacity reconstructing raw, noisy sequence details that may not carry functional biological meaning.

But a promising, more efficient alternative is emerging: Joint-Embedding Predictive Architecture (JEPA)

Originally introduced by Yann LeCun for computer vision, JEPA is a non-generative, self-supervised learning (SSL) framework. Instead of predicting raw inputs, it operates as a world model that predicts abstract semantic embeddings in latent space.

Recently, the JEPA framework (and its more efficient LeJEPA variant) has been adapted into the biological sciences to develop performing foundation models and to improve on already existing ones.

It's interesting how each adaptation modified and tailored JEPA to suit its specific biological domain, whether by experimenting with different backbones or complementing the objective with other loss terms.

For example, JEPA-DNA and ProteinJEPA used JEPA as a continual pre-training framework to enhance existing foundation models without training from scratch, while Cell-JEPA and JEPA-DNA employed a hybrid objective that combines the JEPA loss with a traditional language modeling loss.

The article below provides an overview of these implementations, along with others that came out this year. As always, your thoughts and feedback are welcome and highly appreciated!

Link to the article is in the first comment 👇

👁 Image

reacted to pankajpandey-dev's post with 🔥 29 days ago

view post

🇮🇳 Qwen3-4B Hindi Instruct v2 — a Hindi LLM that runs on your own machine
Most strong Hindi-capable models are either huge or cloud-only. I wanted one that's small enough to run locally but actually follows instructions in Hindi — so I fine-tuned Qwen3-4B on 10K Hindi instruction pairs and shipped it with a full GGUF quant ladder.
✅ Fine-tune (16-bit): huggingface.co/pankajpandey-dev/Qwen3-4B-Hindi-Instruct-v2
✅ GGUF (Q4/Q5/Q8): huggingface.co/pankajpandey-dev/Qwen3-4B-Hindi-Instruct-v2-GGUF
Runs in Ollama, llama.cpp, and LM Studio. The Q4_K_M is just 2.5 GB — fits comfortably on a laptop, CPU or GPU.
Part of my Hindi LLM Series — building openly-licensed Indic models for local and edge use. More coming (Gemma next). Feedback welcome 🙏
#Hindi #IndicNLP #GGUF #LocalLLM #Qwen

👁 Image

liked 2 models 30 days ago

Fill-Mask • 6B • Updated 25 days ago • 1.46M • 18

0.2B • Updated 25 days ago • 271k • 42

👁 Image

upvoted a paper about 1 month ago

Paper • 2605.08044 • Published May 8 • 12

👁 Image

liked a model 2 months ago

1B • Updated Apr 23 • 43 • 4

👁 Image

updated a Space 3 months ago

DeepSite Project

🛠

Explore and download Chenopodium genome assemblies

👁 Image

published a Space 3 months ago

DeepSite Project

🛠

Explore and download Chenopodium genome assemblies

👁 Image

liked a model 4 months ago

1.0B • Updated Mar 10 • 158 • 4

👁 Image

liked a dataset 5 months ago

Viewer • Updated 18 days ago • 1.67M • 29.7k • 235

👁 Image

liked a model 5 months ago

Updated Jan 27 • 105

👁 Image

upvoted a collection 5 months ago

Collection of AlphaGenome models. • 5 items • Updated Mar 12 • 44

👁 Image

New activity in scarysnake/outfitter-advice 6 months ago

[bot] Conversion to Parquet

#1 opened 11 months ago by

👁 Image

parquet-converter

👁 Image

updated a dataset 6 months ago

Viewer • Updated Jan 11 • 12 • 16 • 1

👁 Image

upvoted a collection 6 months ago

Best Datasets for Arabic Speech Tasks • 22 items • Updated 26 days ago • 21

👁 Image

reacted to MohamedRashad's post with ❤️ 6 months ago

view post

I have update my https://huggingface.co/collections/MohamedRashad/arabic-speech-datasets
with new datasets, making the full audio data more than 3000 hours of good arabic speech.

Feel Free to use it in your new innovations, And happy new year!

👁 Image

liked 3 models 6 months ago

Text Generation • 0.3B • Updated Jan 14 • 15.9k • 1.02k

Text Generation • Updated Jan 24, 2023 • 1.9k • 157

Text Generation • 32B • Updated Dec 31, 2025 • 19.9k • 318

URL: https://huggingface.co/monsoon-nlp/activity/all

⇱ monsoon-nlp (Nick Doiron)

Nick Doiron

AI & ML interests

Recent Activity

Organizations

DeepSite Project

DeepSite Project

[bot] Conversion to Parquet