Address
:
[go:
up one dir
,
main page
]
Include Form
Remove Scripts
Accept Cookies
Show Images
Show Referer
Rotate13
Base64
Strip Meta
Strip Title
Session Cookies
Skip to content
Navigation menu
Search
Powered by Algolia
Search
Log in
Create account
DEV Community
Close
#
moe
Follow
Hide
Posts
Left menu
đź‘‹
Sign in
for the ability to sort posts by
relevant
,
latest
, or
top
.
Right menu
NVIDIA Unveils Nemotron-Labs-3-Puzzle-75B-A9B-NVFP4: A Deployment-Optimized Hybrid MoE LLM
Pneumetron
Pneumetron
Pneumetron
Follow
Jul 13
NVIDIA Unveils Nemotron-Labs-3-Puzzle-75B-A9B-NVFP4: A Deployment-Optimized Hybrid MoE LLM
#
nvidia
#
nemotron3puzzle
#
llm
#
moe
Comments
Add Comment
4 min read
2 TB of Ukrainian Law + DeepSeek V3 860B on GCP: What We'd Get
overthelex
overthelex
overthelex
Follow
Jul 3
2 TB of Ukrainian Law + DeepSeek V3 860B on GCP: What We'd Get
#
deepseekv3
#
moe
#
tpuv5p
#
gcp
Comments
Add Comment
7 min read
Scaling MoE Models with LongCat-2.0: A Deep Dive into 1.6T Parameter Architecture Design
Tamiz Uddin
Tamiz Uddin
Tamiz Uddin
Follow
Jun 30
Scaling MoE Models with LongCat-2.0: A Deep Dive into 1.6T Parameter Architecture Design
#
ai
#
moe
#
models
#
longcat
Comments
Add Comment
3 min read
Step 3.7 Flash is a drop-in — except for one endpoint detail
Creeta
Creeta
Creeta
Follow
Jun 18
Step 3.7 Flash is a drop-in — except for one endpoint detail
#
stepfun
#
step37flash
#
llm
#
moe
1
 reaction
Comments
Add Comment
9 min read
Kimi K2.6 for Local AI in 2026: What VRAM and System RAM You Need to Actually Run the 1T-Parameter MoE Coding Leader
Jovan Chan
Jovan Chan
Jovan Chan
Follow
Jun 12
Kimi K2.6 for Local AI in 2026: What VRAM and System RAM You Need to Actually Run the 1T-Parameter MoE Coding Leader
#
kimik2
#
localllm
#
moe
#
hardwareguide
Comments
Add Comment
6 min read
I built a Rust inference engine that streams MoE expert weights from NVMe SSDs, no GPU required
Randy AP
Randy AP
Randy AP
Follow
May 27
I built a Rust inference engine that streams MoE expert weights from NVMe SSDs, no GPU required
#
ai
#
rust
#
moe
Comments
Add Comment
2 min read
Mistral Large 3: The 675B Open-Weight MoE Model Developer Guide
Jangwook Kim
Jangwook Kim
Jangwook Kim
Follow
May 5
Mistral Large 3: The 675B Open-Weight MoE Model Developer Guide
#
mistral
#
llm
#
opensource
#
moe
Comments
Add Comment
5 min read
ZAYA1-8B: a 760M-active MoE trained on AMD MI300x
Thousand Miles AI
Thousand Miles AI
Thousand Miles AI
Follow
May 22
ZAYA1-8B: a 760M-active MoE trained on AMD MI300x
#
ai
#
llm
#
moe
#
openweights
Comments
Add Comment
5 min read
đź‘‹
Sign in
for the ability to sort posts by
relevant
,
latest
, or
top
.
We're a place where coders share, stay up-to-date and grow their careers.
Log in
Create account