Skip to content
Navigation menu
Search
Powered by Algolia
Search
Log in
Create account
DEV Community
Close
#
pytorch
Follow
Hide
Posts
Left menu
👋
Sign in
for the ability to sort posts by
relevant
,
latest
, or
top
.
Right menu
I Trained a 6.4M-Parameter Transformer From Scratch to Talk About Recipes
Medha
Medha
Medha
Follow
Jul 25
I Trained a 6.4M-Parameter Transformer From Scratch to Talk About Recipes
#
machinelearning
#
pytorch
#
llm
#
python
1
 reaction
Comments
Add Comment
5 min read
Why My Medical AI Took 6.4 Seconds Per Scan and How I Got It to 3.1.
Sowaiba Arshad
Sowaiba Arshad
Sowaiba Arshad
Follow
Jul 25
Why My Medical AI Took 6.4 Seconds Per Scan and How I Got It to 3.1.
#
ai
#
machinelearning
#
pytorch
Comments
2
 comments
6 min read
RoPE: How 2D Rotations Solved Transformer Long-Context
Masih Maafi
Masih Maafi
Masih Maafi
Follow
Jul 22
RoPE: How 2D Rotations Solved Transformer Long-Context
#
python
#
machinelearning
#
ai
#
pytorch
1
 reaction
Comments
Add Comment
4 min read
Running Qwen3 Through the ExecuTorch MLX Delegate: Up to 4.52x Faster on M1 Max
Nariaki Wada
Nariaki Wada
Nariaki Wada
Follow
Jul 22
Running Qwen3 Through the ExecuTorch MLX Delegate: Up to 4.52x Faster on M1 Max
#
python
#
pytorch
#
llm
#
applesilicon
Comments
Add Comment
7 min read
Testing PyTorch 2.13 MPS FlexAttention on M1 Max: Up to 7.83x Faster for Sparse Attention
Nariaki Wada
Nariaki Wada
Nariaki Wada
Follow
Jul 21
Testing PyTorch 2.13 MPS FlexAttention on M1 Max: Up to 7.83x Faster for Sparse Attention
#
python
#
pytorch
#
machinelearning
#
applesilicon
Comments
Add Comment
7 min read
What Does `unsqueeze` Do in PyTorch? (And Why Your Model Keeps Asking For It)
Wesam Khallaf — Author of PyTorch From Ground Up
Wesam Khallaf — Author of PyTorch From Ground Up
Wesam Khallaf — Author of PyTorch From Ground Up
Follow
Jul 19
What Does `unsqueeze` Do in PyTorch? (And Why Your Model Keeps Asking For It)
#
pytorch
#
python
#
beginners
#
deeplearning
Comments
Add Comment
7 min read
PyTorch Broadcasting Explained: The 3 Rules (and the Silent Bug That Bites Everyone)
Wesam Khallaf — Author of PyTorch From Ground Up
Wesam Khallaf — Author of PyTorch From Ground Up
Wesam Khallaf — Author of PyTorch From Ground Up
Follow
Jul 16
PyTorch Broadcasting Explained: The 3 Rules (and the Silent Bug That Bites Everyone)
#
python
#
pytorch
#
broadcasting
#
machinelearning
1
 reaction
Comments
Add Comment
4 min read
Debugging a Python "Memory Leak" That Was Actually a Measurement Bug (ru_maxrss vs VmRSS)
flipslidersand
flipslidersand
flipslidersand
Follow
Jul 9
Debugging a Python "Memory Leak" That Was Actually a Measurement Bug (ru_maxrss vs VmRSS)
#
python
#
debugging
#
pytorch
#
rag
Comments
Add Comment
4 min read
Classifier-free guidance above 7.5 oversaturated our product renders
Elise Moreau
Elise Moreau
Elise Moreau
Follow
Jun 26
Classifier-free guidance above 7.5 oversaturated our product renders
#
machinelearning
#
computervision
#
pytorch
1
 reaction
Comments
Add Comment
4 min read
Using the channels-last memory format reduced the latency of our conversation backbone by 22%
Elise Moreau
Elise Moreau
Elise Moreau
Follow
Jun 24
Using the channels-last memory format reduced the latency of our conversation backbone by 22%
#
pytorch
#
computervision
#
machinelearning
#
mlops
1
 reaction
Comments
Add Comment
4 min read
The SDXL VAE overflow that decoded black images in fp16
Elise Moreau
Elise Moreau
Elise Moreau
Follow
Jun 23
The SDXL VAE overflow that decoded black images in fp16
#
pytorch
#
computervision
#
machinelearning
#
mlops
1
 reaction
Comments
Add Comment
4 min read
Data Science Workload: Giới hạn RAM trên Dell Pro Max 14 MC14250
Review Laptop
Review Laptop
Review Laptop
Follow
Jun 23
Data Science Workload: Giới hạn RAM trên Dell Pro Max 14 MC14250
#
dellpromax14
#
datascience
#
pytorch
#
jupyter
Comments
Add Comment
3 min read
The seam our tiled upscaler left on every 4K product render
Elise Moreau
Elise Moreau
Elise Moreau
Follow
Jun 19
The seam our tiled upscaler left on every 4K product render
#
mlops
#
computervision
#
pytorch
#
machinelearning
Comments
Add Comment
4 min read
Perplexity held flat after INT4. Task accuracy dropped 7 points.
Marcus Chen
Marcus Chen
Marcus Chen
Follow
Jun 19
Perplexity held flat after INT4. Task accuracy dropped 7 points.
#
machinelearning
#
llm
#
mlops
#
pytorch
Comments
Add Comment
4 min read
Developer Take On: A High-Resolution Neural Cellular Automata
Kelvin Kariuki
Kelvin Kariuki
Kelvin Kariuki
Follow
Jun 17
Developer Take On: A High-Resolution Neural Cellular Automata
#
ai
#
machinelearning
#
pytorch
#
neuralnetworks
Comments
Add Comment
4 min read
👋
Sign in
for the ability to sort posts by
relevant
,
latest
, or
top
.
We're a place where coders share, stay up-to-date and grow their careers.
Log in
Create account