Upgrade to Pro
— share decks privately, control downloads, hide ads and more …
Speaker Deck
Sign up for free
Menu
Search
Features
All features
Private URLs
Password Protection
Custom URLS
Scheduled publishing
Remove Branding
Restrict embedding
Deck Collections
Notes
Features
All features
Private URLs
Password Protection
Custom URLS
Scheduled publishing
Remove Branding
Restrict embedding
Deck Collections
Notes
Explore
Featured decks
Featured speakers
Programming
Technology
Storyboards
Explore
Featured decks
Featured speakers
Programming
Technology
Storyboards
Pricing
Search
Sign in
Sign up for free
Linear Algebra at Large Scale
Search
Elizabeth Ramirez
April 27, 2018
Science
930
7
Share
Embed
Copy iframe code
Copy JS code
Copy link
Start on current slide
Linear Algebra at Large Scale
Elizabeth Ramirez
April 27, 2018
More Decks by Elizabeth Ramirez
See All by Elizabeth Ramirez
Maritime Transportation from Space: The most important industry you know nothing about.
eramirem
0
61
LADL-Code Mesh V
eramirem
0
230
Transition Matrix Estimation in High Dimensional Time Series.
eramirem
0
290
The Linear Algebra of Deep Learning
eramirem
2
770
Linear Algebra for FE Developers
eramirem
1
650
Top 10: Los mejores algoritmos del Siglo XX
eramirem
0
510
Numerical Analysis for Orbit Propagation
eramirem
0
300
A New Approach to Linear Filtering and Prediction Problems
eramirem
0
1.6k
Kalman Filters for non-rocket science - PyCon 2016
eramirem
2
420
Other Decks in Science
See All in Science
2026 Introduction to University Math 01
kanaya
0
140
Inside the Mind of an LLM
baggiponte
0
330
データベース08: 実体関連モデルとは?
trycycle
PRO
0
1.6k
From Prediction to Understanding: Causal Discovery for Data Science and AI Applications
sshimizu2006
0
300
CVPR2026_VGGTとその仲間たち
mickey_0226
0
1.1k
J-STAGE全文XML登載必須化について
xspa2012
0
1.5k
データベース14: B+木 & ハッシュ索引
trycycle
PRO
0
920
(CVPR2026) Back to Basics: Let Denoising Generative Models Denoise
shumpei777
0
350
1. CPC理論の展開と集合的知能モデル(JSAI2026 KS-27 集合的予測符号化と新たな知性の時代)
hayashiyus884
1
360
「念のためのログ保存」を組織全体でやめるためのポリシーと仕組み作り
i2tsuki
4
370
データベース01: データベースを使わない世界
trycycle
PRO
1
1.5k
Conversation is the New Dashboard: 属人性を排除する第4世代BIツールの勢力図
shomaekawa
1
650
Featured
See All Featured
Creating an realtime collaboration tool: Agile Flush - .NET Oxford
marcduiker
35
2.6k
A designer walks into a library…
pauljervisheath
211
25k
Automating Front-end Workflow
addyosmani
1369
210k
Why You Should Never Use an ORM
jnunemaker
PRO
61
10k
Measuring Dark Social's Impact On Conversion and Attribution
stephenakadiri
2
270
Side Projects
sachag
456
43k
Let's Do A Bunch of Simple Stuff to Make Websites Faster
chriscoyier
508
140k
Agile Actions for Facilitating Distributed Teams - ADO2019
mkilby
0
280
Why Your Marketing Sucks and What You Can Do About It - Sophie Logan
marketingsoph
0
410
First, design no harm
axbom
PRO
2
1.3k
The B2B funnel & how to create a winning content strategy
katarinadahlin
PRO
1
510
RailsConf 2023
tenderlove
30
1.5k
Transcript
Linear Algebra at Large Scale Elizabeth Ramirez @eramirem
Computational Engineer We model complex systems on the planet, like
forestry and agriculture using satellite imagery.
None
Top 10 Algorithms of the 20th Century
Often the most expensive computations in large-scale codes. Curse of
Dimensionality
Linear Systems Nonlinear Systems Machine Learning Deep Learning
Most ubiquitous problem in Scientific Computing and Data Analysis
What solves? Systems of Equations Polynomial Interpolation Linear Least-Squares
What we know? Gaussian Elimination Complexity
HPC Alternative: Iterative Methods General Form
Jacobi Gauss-Seidel
Convergence of Basic Iterative Methods Spectral radius
Krylov Subspaces
Conjugate Gradient Method (CG) i) ii)
Conjugate Gradient (CG)
Bi-conjugate gradient (BiCG) Any linear system
Deep Learning Primitives Weights, inputs, outputs stored in tensors Matrix
Multiplication Convolution Inner Product Transposition Rectified Linear Unit (ReLu)
Matrix Multiplication Fundamental task Naive: Strassen:
Low-Rank Approximation Accelerates matrix multiplication, therefore, accelerates convolution. Requires SVD:
Low-Rank Multiplication:
Single Instruction Multiple Data (SIMD) Data-level parallelism Incompatible with code
designed for sequential processors Instruction set available in commercial CPUs and GPGPUs
Intel® Math Kernel Library (Intel® MKL) Improved Matrix Multiplication Performance
in LAPACK LU decomposition and inverse without pivoting Take advantage of SIMD instruction set In summary: High Performance Linear Algebra
None
References http://www.siam.org/pdf/news/637.pdf https://software.intel.com/en-us/mkl https://software.intel.com/en-us/articles/t ensorflow-optimizations-on-modern-intel-arc hitecture