Upgrade to Pro
— share decks privately, control downloads, hide ads and more …
Speaker Deck
Features
Speaker Deck
PRO
Sign in
Sign up for free
Search
Search
Seeing at the Speed of Thought: Empowering Othe...
Search
Sponsored
·
Your Podcast. Everywhere. Effortlessly.
Share. Educate. Inspire. Entertain. You do you. We'll handle the rest.
→
Greg Goltsov
March 08, 2017
Programming
0
260
Seeing at the Speed of Thought: Empowering Others Through Data Exploration
Talk I gave at Big Data Visualisation Sydney 2017
Greg Goltsov
March 08, 2017
Tweet
Share
More Decks by Greg Goltsov
See All by Greg Goltsov
Beginning ClojureScript: How not to learn a new language
ggoltsov
1
120
Full-stack Data Science: How to be a One-Man Data Team
ggoltsov
2
620
Kuranku - final game presentation
ggoltsov
0
300
Scalable agent-based simulations
ggoltsov
1
240
Procedural City Generator - Honours Presentation
ggoltsov
2
1.6k
Extracting the Meaning: Painless processing and analysis of image data with Fiji and Ruby
ggoltsov
0
180
Ninja Code
ggoltsov
1
2.4k
Other Decks in Programming
See All in Programming
AIコードレビューの導入・運用と AI駆動開発における「AI4QA」の取り組みについて
hagevvashi
0
550
我々はなぜ「層」を分けるのか〜「関心の分離」と「抽象化」で手に入れる変更に強いシンプルな設計〜 #phperkaigi / PHPerKaigi 2026
shogogg
2
370
GoのDB アクセスにおける 「型安全」と「柔軟性」の両立 - Bob という選択肢
tak848
0
270
テレメトリーシグナルが導くパフォーマンス最適化 / Performance Optimization Driven by Telemetry Signals
seike460
PRO
2
150
OTP を自動で入力する裏技
megabitsenmzq
0
120
Windows on Ryzen and I
seosoft
0
380
ベクトル検索のフィルタを用いた機械学習モデルとの統合 / python-meetup-fukuoka-06-vector-attr
monochromegane
2
520
エンジニアの「手元の自動化」を加速するn8n 2026.02.27
symy2co
0
180
Symfony + NelmioApiDocBundle を使った スキーマ駆動開発 / Schema Driven Development with NelmioApiDocBundle
okashoi
0
220
Java 21/25 Virtual Threads 소개
debop
0
270
20260313 - Grafana & Friends Taipei #1 - Kubernetes v1.36 的開發雜記:那些困在 Alpha 加護病房太久的 Metrics
tico88612
0
230
最初からAWS CDKで技術検証してもいいんじゃない?
akihisaikeda
4
170
Featured
See All Featured
The Limits of Empathy - UXLibs8
cassininazir
1
270
First, design no harm
axbom
PRO
2
1.1k
Intergalactic Javascript Robots from Outer Space
tanoku
273
27k
Become a Pro
speakerdeck
PRO
31
5.9k
Information Architects: The Missing Link in Design Systems
soysaucechin
0
840
Build The Right Thing And Hit Your Dates
maggiecrowley
39
3.1k
The Language of Interfaces
destraynor
162
26k
Noah Learner - AI + Me: how we built a GSC Bulk Export data pipeline
techseoconnect
PRO
0
150
Site-Speed That Sticks
csswizardry
13
1.1k
The Anti-SEO Checklist Checklist. Pubcon Cyber Week
ryanjones
0
100
Keith and Marios Guide to Fast Websites
keithpitt
413
23k
Git: the NoSQL Database
bkeepers
PRO
432
67k
Transcript
Seeing at the speed of thought Empowering others through data
exploration Greg Goltsov Senior Data Engineer @gregoltsov www.gregory.goltsov.info (will have link to slides)
Seeing at the speed of thought Empowering others through data
exploration
Seeing at the speed of thought Empowering others through data
exploration
Seeing at the speed of thought Empowering others through data
exploration yourself
Seeing at the speed of thought Empowering others through data
exploration yourself your team
Seeing at the speed of thought Empowering others through data
exploration yourself your team your company
Touch Surgery Built marketing/sales dashboards for Fortune 10 companies Built
educational dashboards for 4 of the top 10 world-rated medical universities All from scratch
Appear Here World’s biggest online marketplace for retail spaces Internal
recommendation system Highly visual debug interface for non-tech people
Southern Cross Austereo Modernising the data pipeline Spearheading data-driven culture
throughout the company Datasets covering 80% Australians weekly
BI/DW tools
BI/DW tools
Remove barriers Make feedback fast Remove yourself
Remove barriers
Remove barriers Catalogued datasets with one-line import in Python Messy
dataset in PDFs
Remove barriers Dashboard with right filters, Excel export “Can you
run a query?”
Remove barriers. Foster curiosity.
Make feedback fast
Make feedback fast Found a new trend via tinkering “Tomorrow
I’ll see results of the batch job”
Make feedback fast “Check the dash in 15 mins” “I
put your request into the backlog”
Make feedback fast. Let people tinker.
Remove yourself
Remove yourself Data pipeline + products Ad-hoc
None
None
Remove yourself. Don’t stand in the way.
Remove barriers Make feedback fast Remove yourself
The goal is to turn data into information, and information
into insight. – Carly Fiorina, former HP CEO
Insight Information Data
Insight Information Data Value ↑ Abundance
Insight Information Data Fraud Access pattern Logs
Insight Information Data Key influencers MOM trends Tweets
Ad-hoc queries Data pipeline Fast to develop Every query gets
thrown away after Upfront investment Every integration builds foundations
Visualise your ETL. Augment your Data Warehouses with Data Lakes.
None
Extract Transform Load Sources Data Warehouse
Extract Transform Load Sources Data Warehouse Data Insight Time
Volume Variety Velocity "3D Data Management: Controlling Data Volume, Velocity
and Variety”, Gartner Inc. 2001
Volume Variety Velocity "3D Data Management: Controlling Data Volume, Velocity
and Variety”, Gartner Inc. 2001
Analysis of Unstructured Data: Applications of Text Analytics and Sentiment
Mining ~80% of all data is unstructured
~80% of your data is unstructured
http://www.ft.com/cms/s/0/de15414e-ebad-11e1-985a-00144feab49a.html#axzz2F3CM6G7g “Making sense of unstructured data isn’t about technology, it’s
a business challenge”
Aberdeen Group research Don’t use unstructured data Use unstructured data
Happy with the ability to share data 18% 60% Pleased with the accessibility 20% 50%
Volume Variety Velocity Machine learning "3D Data Management: Controlling Data
Volume, Velocity and Variety”, Gartner Inc. 2001
Ingest quickly Real-time schema-on- read exploration Push vetted insights into
DW/BI Example: Spark, AWS Athena, Microsoft’s PowerBI
Collect Store Process/ Analyse Sources Data Warehouse Data Insight Insight
Time
Collect Store Process/ Analyse
Collect Store Process/ Analyse
Collect Store Process/ Analyse
None
Look at data. A lot.
Look at data. A lot. http:/ /www.forbes.com/sites/gilpress/2016/03/23/data- preparation-most-time-consuming-least-enjoyable- data-science-task-survey-says
None
None
Scale computation and storage separately Go from non-trivial data to
dashboard in minutes Spark is 20-100x faster than MapReduce Turnkey solution: www.databricks.com OSS: Apache Zeppelin on AWS EMR Spark
We made it! Now what?
We made it! Now what? Human scale.
AirBnB Scaling Tribal Knowledge
AirBnB Scaling Tribal Knowledge
AirBnB Scaling Tribal Knowledge
AirBnB Scaling Tribal Knowledge
AirBnB Scaling Tribal Knowledge
None
THANK YOU Speaker Name: Greg Goltsov Email:
[email protected]
Organized by
UNICOM Trainings & Seminars Pvt. Ltd.
[email protected]
http://www.unicomlearning.com/2017/Big_Data_Visualization_Summit_Sydney