Co-Op, Data Extraction
Lila Sciences - Cambridge, MA USA
Posted Jun 11, 2026
Benefits
- Parental leave
- Not verified
- Non-birth-parent leave
- Not verified
- Family-building benefits
-
- Fertility benefits: Not verified
- Adoption assistance: Not verified
- Surrogacy assistance: Not verified
- Mental health support
- Not verified
- Relocation assistance
- Not verified
- Childcare support
- Not verified
- Learning budget
- Not verified
- Verification
- Not verified
- Salary
- Not verified
- 401(k) match
- Not verified
Was this benefit information wrong? Tell us.
Market context
- U.S. role benchmark (BLS OEWS)
- $111,944 U.S. median for this role
- Projected growth (BLS Employment Projections)
- +13.7% - Much faster than average
Matched to SOC 15-1252 - Data and ML aggregate by role bucket.
Source: U.S. Bureau of Labor Statistics, OEWS, May 2024 and Employment Projections, 2024-2034.
Role
Schedule
- Shift type
- Not verified
- Weekend work
- Not verified
Application
- Cover letter
- Not verified
- Assessment
- Not verified
- Deadline
- Not stated
Where they hire
State eligibility is not yet verified.
About this role
Co-Op, Data Extraction Cambridge, MA USA Your Impact at LILA Lila Sciences builds AI systems that accelerate discovery across the physical and life sciences. Within Physical Sciences AI, our team works on turning unstructured scientific knowledge (e.g., literature, patents, technical reports) into structured signals that power downstream Lila applications. As a Data Extraction Co-Op, you will work alongside research scientists and engineers on a focused sub-problem in this stack. You will get hands-on experience fine-tuning and evaluating extraction models, building pipelines for messy real-world data, and shipping work that flows into production systems. What You'll Be Building - Contribute to AI systems that extract and structure knowledge from scientific literature and patents, focused on a well-defined sub-problem - Fine-tune and evaluate language, multimodal, or specialized models for data extraction, with mentor guidance - Build and test pipelines that structure unstructured scientific data across text, tables, and visuals - Run extraction pipelines, analyze results, and document findings clearly - Share your work through a team presentation, write-up, or contribution to a publication or open-source project What You'll Need to Succeed - Pursuing a Bachelor's, Master's, or PhD in Computer Science, Chemistry, Materials Science, or a related field - Solid foundation in machine learning fundamentals and Python - Familiarity with NLP or computer vision concepts - Curiosity about scientific data and willingness to learn quickly in a research setting Bonus Points For - Coursework or projects involving multimodal models or document understanding (OCR, table/figure extraction) - Experience working with messy, real-world datasets
Read the full description at job-boards.greenhouse.io. FewerJobs shows a preview and links to the original posting.
Apply link not verified; last-live date unavailable.
What verified means
Verified means a displayed claim has field-level provenance to a source FewerJobs pulled: a government or employer source, or the original job posting. Posting-sourced facts are employer-stated and are labeled separately from government records.
Related jobs
-
Senior Data Scientist
Cooper Companies (The) - San Ramon, CA, United States
-
Data Scientist
Accendra Health INC - Philadelphia, PA
-
Data Analyst - Big Data and Analytics Platform
ICF International INC - Reston, VA
-
Staff Data Scientist
Heartflow INC - San Francisco, California