Ancestry
Data Science Engineering, Intern
About Ancestry:
When you join Ancestry, you join a human-centered company where every person’s story is important. We believe that by discovering the struggles and triumphs of our past, we can foster deeper bonds and more meaningful connections among families and communities. Our talented team of scientists, engineers, genealogists, historians, and storytellers is dedicated to empowering customers around the world from all backgrounds on their journeys of personal discovery.
With more than 27+ billion digitized global historical records, 100 million family trees, and 18 million people in our growing AncestryDNA database, Ancestry helps customers discover their family story and gain a new level of understanding about their lives. Passionate about dedicating your work to enriching people’s lives? You belong at Ancestry.
Ancestry is looking for an exceptional, passionate, and highly motivated Data Science Engineering Intern to join our Data Science Engineering team. The DSE team develops products and tools to empower data science and ML use cases. DSE provides data pipelines and data APIs to data scientists. As a DSE Intern, you will gain hands-on experience with scoping, development and operationalization aspects of data pipelines for feature generation to enable various machine learning use cases. As a Data Science Engineer Intern, you will create data science pipelines and RESTful APIs for both internal and external customers. You will also interact with the Data Science team in understanding business requirements and building features and products to unlock new possibilities with data.
What You'll Do:
- Work on new and existing features to aid new machine learning models
- Work on the deployment of data science models in a production pipeline
- Create tooling and automation to enable data flow and model performance monitoring
- Write new data pipeline jobs to retrieve, process and validate training data
- Develop new features on web services
- Develop new data pipelines to generate and validate training data feature sets
- Learn various phases involved in development of a project from prototyping to productionalizing
- Help create infrastructure using our Infrastructure as Code (IAC) framework
- Create tests to measure performance and monitor behavior of ML pipelines
Who You Are:
- Must be enrolled in a College or University in the US and graduating after Summer of 2021
- Must have a strong foundation in Java 8/11 and Python
- Familiarity with Linux shell (Bash, csh etc.)
- Familiarity with SQL
- Strong written and verbal communication skills
- Curiosity and “Go-Getter” attitude
Nice to Have:
- Experience with Big Data technologies (Spark, kafka, Scala)
- Familiarity with AWS (Lambda, Kinesis, SNS, SQS etc.)
- Familiarity with machine learning technologies (scikit-learn, tensorflow, pytorch)
It takes customer obsession to pioneer and focus on what matters most, to tackle challenges, break through boundaries to create change, to stay on the cutting edge, and pioneer relentlessly. We are looking for the next generation of talent to drive the business, grow our people, and put the customer before all else. Our culture empowers journeys of personal discovery, giving you the freedom to participate in impactful projects and explore the intersections of our businesses: technology, data science, Health, DNA science, product, design, marketing, legal, and more. We’re looking for innovative minds, diverse ideas, and for students interested in connecting personal discoveries and technology. Our Interns enjoy mentorship and experience challenging work while receiving fantastic pay, fully-paid temporary housing, and having a fun captivating experience—we have it all. Oh, and did we mention the possibility of full-time employment once you graduate?
IND2
#LI-SJ2
Additional Information:
Ancestry is an Equal Opportunity Employer that makes employment decisions without regard to race, color, religious creed, national origin, ancestry, sex, pregnancy, sexual orientation, gender, gender identity, gender expression, age, mental or physical disability, medical condition, military or veteran status, citizenship, marital status, genetic information, or any other characteristic protected by applicable law. In addition, Ancestry will provide reasonable accommodations for qualified individuals with disabilities.
All job offers are contingent on a background check screen that complies with applicable law. For San Francisco office candidates, pursuant to the San Francisco Fair Chance Ordinance, Ancestry will consider for employment qualified applicants with arrest and conviction records.
Ancestry is not accepting unsolicited assistance from search firms for this employment opportunity. All resumes submitted by search firms to any employee at Ancestry via-email, the Internet or in any form and/or method without a valid written search agreement in place for this position will be deemed the sole property of Ancestry. No fee will be paid in the event the candidate is hired by Ancestry as a result of the referral or through other means.
- Founded
- Founded 1996
- Employees
- 500+ employees
- Industry
- Internet Software & Services
- Total raised
- $330M raised