Showing posts with label Data Engineer. Show all posts
Showing posts with label Data Engineer. Show all posts

Wednesday, April 1, 2015

[Prezi] Data Engineer - San Francisco

Challenges

Build the data infrastructure and frameworks that will be used to find correlations across Marketing and Sales that will form testable hypotheses and upon which experiments will be judged.
Conduct data engineering and analysis, and while partnering with our data analyst, arrive at conclusions and make actionable business recommendations based on your research and investigations. You’ll need to work smart, fast, and cross-functionally.

Responsibilities

  • Partner with our San Francisco-based data analyst to test, build, automate, and maintain datasets for our Marketing and Sales teams that will be used to measure the success of new campaigns and product features

  • Turn business questions into testable hypotheses and hypotheses into scalable and definitive experiments (ie: what does it take to convert a free user to paying?). These experiments will inform the highest-level business decisions at Prezi.

  • Triage problems with our campaigns and products by investigating user engagement logs and datasets

Requirements

  • 3-5 years of experience working with Big Data (Internet, banking, telecommunication, or cloud SaaS)

  • Experience in data transformation: SQL, Pig, Python, and Shell scripting

  • Experience in basic statistical methodology: eg. clustering and testing, predictive modelling (eg. decision trees, regressions)

  • Know at least one statistical software in depth: R, Rapidminer, SPSS, SAS, MATLAB etc.


We work with:

  • Pig, Bash, Python, SQL, R

  • Hadoop system

  • Amazon Web Services: EC2, EMR, Redshift

Why work at Prezi?

  • Continuing opportunities to learn new technologies

  • Be empowered to affect Prezi’s business strategy

  • Immediate impact - your work will be seen by more than 50 million users as soon as you deploy

  • <p
  • Office in downtown San Francisco

  • International work environment and opportunities to travel internationally

  • Creative, fast-paced, flexible, fun, and challenging environment

  • Competitive compensation and benefit package including equity

Thursday, November 27, 2014

Senior Data Engineer | Centro

Description

ABOUT THE PRODUCT AND ENGINEERING TEAM

 

Centro is building the world’s most respected product and engineering organization.  This is a bold claim, and one we do not make lightly.  Of course, we can back it up:  Ad Tech is one of the hottest technology sectors today, with the worlds most distinguished product and engineering organizations vying to be the first to truly and forever disrupt a $200 billion industry. 

 

Of course, delivering technology that changes the way an industry does business requires a massive product, design, and engineering effort.  To do our heavy lifting, we attract talent from an array of disciplines to work in highly-collaborative, agile, cross-functional teams.   Together we tackle the complexity of scale that comes from performing 30 billion transactions per day. We innovate with machine learning and Big Data solutions to deliver optimized campaign performance and recommendations. We create elegant user interfaces that unite all the functions of our platform into one, seamless, intuitive, and delightful experience.

 

Together, we are winning advertising tech, will you join us?

 

ABOUT THE TEAM

The Data Ops team is responsible for transforming Centro's data into actionable information across the organization. This ranges from daily performance reports to real-time historical analysis. Currently we use the Pentaho suite as our business intelligence platform and PostgreSQL for our database platform. We believe in Agile software development and use Scrum as our project management methodology. 'Inspect and Adapt' is more than just a catch-phrase to us. We are willing to have every problem under the sun exactly once, in exchange for never having the same problem twice.

 

ABOUT THE ROLE

Centro is looking for talented, passionate Data Warehouse Engineer with the skill and desire to contribute to our business intelligence team.    Centro’s technology focuses on improving streamlining the entire digital media process, from research thru trafficking, with focus on each step in between.  We are seeking talented Data Warehouse Engineer that can help transform the data created from these activities into actionable information.  We aim to derive meaning from our data that enables us to run our business better, and also equips our clients to advertise smarter.   If you have experience working with large data systems, then we would love to talk with you.  If you are the type of person who comes to work every day expecting to learn, contribute, teach, and have fun, then we think you will fit right in.  

CORE RESPONSIBILITIES

As part of our business intelligence team you are going to collaborate to solve real-world business problems that Centro is tackling.  You'll be working with all parts of the BI stack.  This ranges from design, implementation and support of our data warehouse all the way through how our end users consume information we generate.  You'll work closely with business users, data analysts and software engineers to design and implement data integration flows into the enterprise warehouse.

  • A qualified applicant should be familiar with a variety of the data warehousing concepts and practices. 
  • At a basic level, you can translate business intelligence needs into a system design and then implement it. 
  • You have the capabilities to extend a given BI tool's out-of-box functionalities to provide better user experiences, if that is called for. 
  • You enjoy working directly with the people that use these reports and know how to manage their expectations. 
  • We expect that you will contribute to the culture of learning and continuous improvement on the team. 
  • Also, if you don't already have it, that you will build strong business domain knowledge related to online advertising, campaign planning and execution, ad serving technologies and related topics.

QUALIFICATIONS

  • You have 5-10 years of end-to-end experience with data warehousing and BI systems (including data modeling, ETL architectures and OLAP). 
  • You've designed, built, and supported data warehouses.
  • You have a strong understanding of dimensional modeling and other data warehouse techniques.
  • You have the ability to quickly understand business requirements and transform them into data models.
  • You have experience developing and implementing ETL processes, and dealing with performance and scaling issues.
  • You have strong database and SQL skills. (We use PostGreSQL, but experience with similar database platforms works for us.)
  • You have experience in scripting languages (perl, python, javascript, etc)
  • You are knowledgeable in Java (We use Pentaho as our BI platform)
  • You have expertise in job scheduling, data quality, metadata management and monitoring systems.
  • You have excellent written and verbal communication skills, with an ability to express business concepts in technical terms (and vice-versa).
  • You are able to help troubleshoot problems and resolve issues.

Bonus items:

  • You have experience with Ruby
  • You have experience with the Hadoop eco-system (HDFC, MapReduce, Hive, Pig, etc).

 

EDUCATION

  • No degree can substitute for aptitude, attitude, and integrity so we have no such requirement.  Bachelor’s degrees in business, economics or other related field are nice to have.  An MBA, or other advanced degree related to position, is a strong bonus.

 

ABOUT CENTRO

 

Centro provides unified, cloud-based software to simplify digital media operations. Our media management platform streamlines and scales ad buys across all channels, accessing both guaranteed and biddable inventory, to achieve any campaign objective. Our holistic approach gives marketers a single system of record to fulfill their research, planning, buying, optimization and reporting needs. Since 2001, Centro has successfully planned and executed more than 100,000 national and local campaigns across all digital display platforms and ad format types. Headquartered in Chicago with 33 offices in North America, Centro's success and commitment to culture has led to many accolades, including #1 on Crain's Best Places to Work in Chicago in 2011, 2012, 2013 and 2014.

 

***Centro is an Equal Opportunity Employer and does not discriminate against any employee or applicant on the basis of race, gender, age, disability or any other basis protected under the law.***

Tuesday, November 4, 2014

[Prezi] Data Engineer - Budapest

Prezi is the zooming presentation software that uses an open canvas instead of traditional slides to help people explore ideas, collaborate more effectively, and create visually dynamic presentations.  Founded in 2009, and with offices in San Francisco and Budapest, Prezi provides its users a visually engaging, personalized, way to express their ideas anytime, anywhere. The company’s vision extends well authoring software alone into becoming the inspiration and enabler of world-changing ideas for people, organizations and businesses.

Prezi has enjoyed explosive growth and developed a rapid following of passionate users.  More than 40 million people from over 190 countries use Prezi from their desktops, browsers, iPads and iPhones.

Prezi is rapidly adding new users each month, and more than 1 Prezi is created every second.  The company  has over 150 employees, and is backed by premier investors, including Accel Partners, Sunstone Capital (based in Copenhagen), and TED.


We are looking for talented software engineers to join our Data Infrastructure Team. 


Responsibilities: 

- Participate in building a big data infrastructure as a service for Prezi.

We believe that data analytics should be easy for both technical and non-technical people and the tools we work on aim to make this possible.

- Keep petabyte-scale data flowing through our pipeline.

We have hundreds of data-analytics jobs running every day. Flowkeeper, our home grown data platform takes care of all data flows running correctly and reports arriving on time.

- Participate in the development and operation of an ETL solution with such complexity.

- Develop and maintain new data tools to ensure that Prezi gets more and more data driven. 


We work with:

- JavaScript, Python, Go, Bash

- Amazon Web Services: EC2, EMR, Redshift

- SQL, Hadoop, MapReduce


Required skills: 

- Enthusiasm for DevOps: we write it, we run it!

- A solid background in software engineering.

- Get it done behaviour: you are smart and quick with a focus on delivery. 

- Unix skills are a big plus. 

- Brave enough to try new technologies. You will love to work with us if you are not tied to a single stack. 


What we offer you:

- Commit on your first day

- Be empowered to affect your team's strategy 

- Immediate impact - your work will be seen by more than 40 million users as
soon as you deploy

- Have end-to-end responsibility

- Participate in fellowship trips and work in our San Francisco office

- Paid trips to professional conferences anywhere in the world

- Flexible working hours in our beautiful office environment or occasionally from home

- Enjoy free quality food in our bistro

- Receive stock options to become your own employer

- Be yourself - we are proud to be colourful!

Wednesday, October 29, 2014

Data Engineer job at StyleSeat - AngelList

Details
Skills
Python, Data Analysis, Big Data, Databases
Location
San Francisco
Compensation
Full Time
$120K – $140K Salary
0.1% – 0.2% Equity

Thursday, October 9, 2014

Data Engineer/ETL Specialist | Jobs at Airbnb

No global movement springs from individuals. It takes an entire team united behind something big. Together, we work hard, we laugh a lot, we brainstorm nonstop, we use hundreds of Post-Its a week, and we give the best high-fives in town. Learn more about how Airbnb works:  https://www.airbnb.com/info/how_it_works

Airbnb is seeking an experienced Data Engineer/ETL specialist to join the data science team. This person would contribute to the vision for data infrastructure and business intelligence tools, work with engineers and data scientists to optimize logging, and establish best practices for table schemas and data storage. 

Why is this important to Airbnb?

The pace and quality of our decision-making is dependent upon data availability and reliability. We have a data infrastructure team building cutting-edge tools and a data science team digging deep into insights that can drive decisions; Data Engineers are the glue between these two teams. 

Experience

  • Expertise in building and maintaining reliable ETL jobs
  • Proven ability to work with varied forms of data infrastructure, including relational databases (e.g. SQL), Mapreduce/hadoop, and column store (e.g. Vertica/Redshift)
  • Expert knowledge of SQL/Hive, Pig/Cascading a plus
  • Strong CS fundamentals

Responsibilities

  • Work closely with data scientists and engineers to design and maintain scalable data models and pipelines
  • Develop and maintain cross-platform ETL processes
  • Optimize table schemata based on usage patterns
  • Lead development of architecture and standards for a business metric warehouse
  • Work with data scientists to develop a scalable data visualization platform
  • Help data scientists optimize productionized Pig and Hive queries
  • Implement systems for tracking data quality and consistency

Benefits

  • Stock
  • $2000 yearly employee travel coupon
  • Competitive salaries
  • Paid time off
  • Medical, dental, & vision insurance
  • Life & disability coverage
  • 401K
  • Flexible Spending Accounts
  • Apple equipment
  • Daily breakfast, lunch, and dinner
  • Weekly happy hour

Apply Now

Monday, September 1, 2014

Data Engineer | About Pinterest

Data Engineer

Location

San Francisco

Description

We’re looking for talented data engineers to help us build the next generation of data platforms and products. You’ll work on some of the most interesting big data technologies (Hadoop, Storm, Kafka, Spark, Redshift) in the world. If you’re passionate about big data and data-driven engineering, curious about any problem where data can be applied, and love scaling multi-petabyte datasets, we’d love to hear from you!

 

Responsibilities

  • Design and build a scalable data platform to store, process and query petabytes of data
  • Design and build Pinterest’s internal analytics product and prototype new user-facing data products
  • Apply data-mining and machine learning techniques to large structured and unstructured datasets
  • Diagnose and fix complex distributed systems problems
  • Participate in design and code reviews and help improve our engineering practices

 

Requirements

  • Extensive experience programming in Java or C++
  • Experience programming in a scripting language like Python or Ruby
  • MS or Ph.D. in Computer Science or equivalent work experience
  • Excellent problem solver and communicator, and can document your work
  • Experience with one ore more distributed systems like Hadoop, HBase, Storm, Kafka, Spark or Redshift
  • Background in one or more areas like databases, data processing, distributed systems, statistics or machine learning