Fluid Reset
February 
6th
 at 
6:00pm
RSVP
Text goes here
X

12:00 PM

Schedule Element

Clear your calendar - It's going down! Schedule Blocks kicks off on May 20th, and you're invited to take part in the festivities. Splash HQ (122 W 26th St) is our meeting spot for a night of fun and excitement. Come one, come all, bring a guest, and hang loose. This is going to be epic!

Jet Tech Big Data Meetup

Come and learn about:

• the evolution of our data lake

• how we built end-to-end observability for our data lake

• how we used Spark to make our data pipelines simpler, more robust, and responsive to change 

• and how we've made data science at scale successful. 

Wednesday
, 
February 
06
 at 
6:00pm
RSVP
Text goes here
X

ABOUT

Our Business Intelligence & Data Platform team consolidates input from various source systems into one data-lake and exposes it for analytical purposes and business decision support​.


We listen to everyone’s data, look at it to see if anything is of interest to a party, sift information out, and direct it to the appropriate party.


Our team essentially helps analyze data from our source systems and make it easy for other humans to understand. 

 

SCHEDULE

Doors open at 6:00 p.m., and we've got four talks lined up from 6:30 p.m. to 7:50 p.m. 

6:00 PM

Check-In +  Networking

First things first: Make sure to RSVP here on Splash, so our building's security team has your name on their list.

 

When you arrive day-of, come through the street-side entrance, sign in with security, take the elevator to the 8th floor, and check in with us. 

 

Afterward,  grab food and drinks and network with other visitors before we begin our first talk. 

 

6:30 PM

Evolution of the Data Lake at Jet: Journey So Far and the Road Ahead // Qian Chen 

 Jet started it's Data lake journey in 2016. It was fortunate to be Azure native at the start of the journey. In last 2+ years, our Data lake has gone through significant evolution and benefited from introducing several new capabilities and technologies. And we are not done yet ! In this session, we will discuss evolution so far, lessons learned and the Road ahead. 

 

About the Speaker: Qian Chen is a data engineer at Jet working on both RDBMS and big data stack. Prior to joining Jet he worked in the finance industry.  In his free time, he is trying to pick up playing acoustic guitar.

6:50 pm

Building End-to-End Observability for Jet's Data Lake // Sander Hartlage

An enterprise data lake has dozens of infrastructure components and 100s of data pipelines. At Jet, to run a reliable Data lake operations we have built an end-to-end observability solution using Influx TICK stack. In this session, we will share component driven design of our monitoring and alerting solution and discuss lessons learned along our development journey.

 

About the Speaker: Sander Hartlage is a data engineer at Jet. He drives Jet's Data Lake infrastructure and ensures that it keeps up with cutting edge Big data technologies. Sander has built data products for multiple high velocity technology startups in the New York area for around ten years. Sander enjoys coffee and cats.

7:10 PM

Building Data Processing and ETL pipelines Using Spark 301: Data Source V2, Structured Streaming, Integrating with RDBMS // Kevin Jerrard


Jet extensively uses many of advanced capabilities of Apache Spark. In this session, we will discuss how some of the newer capabilities can be leveraged to make your data pipelines much simpler, robust and responsive to change. We will also discuss how some of the common challenges in integrating distributed processing framework such as Spark with SNP architecture of RDBMS engine such as SQL Server.

 

About the Speaker: Kevin Jerrard is a software engineer working on one of the many data platforms at Jet. After studying Information Science at Cornell University, he worked with a handful of companies in finance and the big data space. In his free time, he enjoys obscure non-fiction and playing tourist in New York museums.

7:30 PM

Enabling Data Science on Jet’s Data Lake // Praful Kava

Successful Data Science at scale requires a tight collaboration between Data Science and Big Data Engineering teams. In this session, Flash team (Big Data Engineering @Jet) will talk about components that you need to get right in order to make Data Science  at scale successful. This segment will also include a quick demo of a cloud hosted notebook solution powered by Big Data technologies such as Spark

 

About the Speaker: Praful Kava leads Jet's Data Lake platform. Over past 20 years, Praful has designed and built Application, Data and ML solutions and high performing engineering teams at companies like IBM, EMC, Dell, CapGemini and McKinsey. Away from work, Praful enjoys keeping up with his 6 year old son.

WHERE & WHEN

Wednesday
, 
February 
06
 at 
6:00pm
RSVP
Text goes here
X

#JETBIGDATA

Share with Friends
Facebook
Twitter
LinkedIn
Link
Powered by
CONTACT THE ORGANIZER
Google   Outlook   iCal   Yahoo

RSVP

How did you hear about this event?
Are you interested in career opportunities at Jet?
Please enter the current industry or field you work in.
Please enter your current job title or role.
Skillset
Please check off what skillsets you possess.
processing image...

Throwing your own event?

Make it awesome with Splash.
Check it out!
Google Icon
Google
Outlook Icon
Outlook
Apple Icon
Apple
Yahoo Icon
Yahoo