Altos Labs Logo

Altos Labs

Staff Software Engineer, Data Curation

Posted 23 Days Ago
Be an Early Applicant
In-Office
9 Locations
Senior level
In-Office
9 Locations
Senior level
The Staff Software Engineer will curate data, build ETL workflows, automate processes, and serve as a liaison between scientific and engineering teams to enhance research data management.
The summary above was generated by AI
Our Mission

Our mission is to restore cell health and resilience through cell rejuvenation to reverse disease, injury, and the disabilities that can occur throughout life.

For more information, see our website at altoslabs.com.

Our Value

Our Single Altos Value: Everyone Owns Achieving Our Inspiring Mission.

Diversity at Altos

We believe that diverse perspectives are foundational to scientific innovation and inquiry. At Altos, exceptional scientists and industry leaders from around the world work together to advance a shared mission. Our intentional focus is on Belonging, so that all employees know that they are valued for their unique perspectives. We are all accountable for sustaining a diverse and inclusive environment.

What You Will Contribute To Altos

Use AI agents to make complex research data FAIR—Findable, Accessible, Interoperable, Reusable—so scientists and product teams can ask richer questions, move faster, and advance discovery. Be part of a team using knowledge and data engineering to enable the transition from manual to LLM‑enabled, agentic data ingestion and curation.  You’ll sit at the intersection of data curation, data and knowledge engineering. Your job is to automate the ingestion and standardization of multi‑source datasets into governed, searchable, analytics-ready assets, and to model the domain knowledge that ties them together. 

Responsibilities

  • Curate and harmonize data. Ingest, profile, clean, normalize, and annotate multi‑modal research datasets (e.g., genomics/transcriptomics, proteomics, imaging/microscopy, CRISPR screens, assay/instrument metadata). Map to controlled vocabularies and standards; manage identifiers, synonyms, and crosswalks.
  • Deliver insights from curated data. Focus on the substance—entities, relationships, and annotations that answer real research and product questions using public domain assets from Ensembl, GEO, PubMed, OMIM, OLS, amongst others. Use pipelines and existing data sources storage pragmatically as tools to deliver content and outcomes.
  • Model knowledge to serve decisions. Capture the concepts and links researchers actually use; keep schemas lightweight and purpose‑built. Leverage OBO Foundry ontologies; define with LinkML; align to the BioLink/Biolink Model; and integrate/serve with platforms such as BioCypher.
  • Quality, governance & AI enablement. Instrument automated checks (tests/expectations), process development to improvement data FAIRification, and LLM‑assisted validations; capture provenance/lineage; codify SOPs; and work to facilitate the migration of processes from manual → automation → agentic (MCP‑integrated) workflows.
  • Serve as a key technical liaison between scientific, data science, and engineering teams, translating complex research needs into scalable and maintainable data solutions.
  • Define and evangelize best practices for data and knowledge engineering across the organization, mentoring junior team members and building reusable, AI-enhanced, enterprise-level components.
Who You AreMinimum Qualifications
  • PhD, Biological Sciences, Computer Science, Software Engineering, or related quantitative field, or equivalent technical experience
  • Candidates should have 8+ years of relevant experience in data curation, ontology/knowledge engineering, or data engineering (or equivalent experience) at a biotechnology company.
  • Mindset: You prioritize data and business objectives over tools; technology is a means to an end.
  • Demonstrably strong Python expertise, particularly in the context of data modeling and processing, with strong skills in both relational (SQL) and graph data stores, and the ability to choose pragmatically between them (e.g., Postgres/Redshift vs. Neo4j/Neptune).
  • Comfortable building pragmatic ETL/ELT workflows in a major cloud (preferably AWS), using orchestration frameworks or AWS-native tools.
  • Active user of AI coding editors such as Cursor, with an active interest in designing and building Model Context Protocol (MCP) applications; motivated to migrate processes from manual → automation → agentic.
  • Mature understanding of data quality, provenance, versioning, and “curation as code,” including hands-on use of testing/validation frameworks.
Preferred Qualifications
  • Experience in basic/exploratory life‑science research across multiple modalities (genomics/transcriptomics, proteomics, imaging/microscopy, screening, model organisms); a user of curated content to achieve research/business outcomes.
  • Experience with a data platform such as lamin.ai.
  • Experience with vector databases and search (e.g., Weaviate, FAISS, pgvector) and AI/LLM frameworks (e.g., LiteLLM, LangChain, LlamaIndex) for retrieval-augmented generation and agent workflows.
  • Experience with OBO Foundry ontologies and modern frameworks such as LinkML, BioLink, and BioCypher, familiarity with graph database technologies (e.g., Neo4j, AWS Neptune) and semantic standards (OWL, RDF, SPARQL).
  • Experience creating lightweight semantic layers and AI/LLM‑assisted curation workflows (LiteLLM, FastMCP).

The salary range for Redwood City, CA:

  • Staff Software Engineer: $221,850 - $300,150

Exact compensation may vary based on skills, experience, and location.


For UK applicants, before submitting your application:

- Please click here to read the Altos Labs EU and UK Applicant Privacy Notice (bit.ly/eu_uk_privacy_notice)
- This Privacy Notice is not a contract, express or implied and it does not set terms or conditions of employment.

Equal Opportunity Employment

We value collaboration and scientific excellence.

We believe that diverse perspectives and a culture of belonging are foundational to scientific innovation and inquiry. At Altos Labs, exceptional scientists and industry leaders from around the world work together to advance a shared mission. Our intentional focus is on Belonging, so that all employees know that they are valued for their unique perspectives. We are all accountable for sustaining an inclusive environment.

Altos Labs provides equal employment opportunities to all employees and applicants for employment, without regard to race, color, religion, age, sex, national origin, disability status, genetics, protected veteran status, sexual orientation, gender identity or expression, or any other characteristic protected by federal, state or local laws. Altos prohibits unlawful discrimination and harassment. This policy applies to all terms and conditions of employment, including recruiting, hiring, placement, promotion, termination, layoff, recall, transfer, leaves of absence, compensation and training.

Thank you for your interest in Altos Labs where we strive for a culture of scientific excellence, learning, and belonging.

Note: Altos Labs will not ask you to download a messaging app for an interview or outlay your own money to get started as an employee. If this sounds like your interaction with people claiming to be with Altos, it is not legitimate and has nothing to do with Altos. Learn more about a common job scam at https://www.linkedin.com/pulse/how-spot-avoid-online-job-scams-biron-clark/

Top Skills

AWS
Aws Neptune
Biocypher
Biolink
Faiss
Langchain
Linkml
Litellm
Llamaindex
Neo4J
Pgvector
Python
SQL
Weaviate

Similar Jobs

58 Minutes Ago
In-Office
8 Locations
Senior level
Senior level
Artificial Intelligence • Big Data • Cloud • Information Technology • Software • Cybersecurity • Data Privacy
Drive sales growth and pipeline management for new and existing accounts in the northwest region, collaborating with Sales Engineers and partners to exceed quotas.
Top Skills: It InfrastructureSaas Security
2 Hours Ago
Hybrid
Oshawa, ON, CAN
Junior
Junior
Automotive • Big Data • Information Technology • Robotics • Software • Transportation • Manufacturing
The Controls Engineer at GM is responsible for developing safety culture, providing technical support for automation and controls, and improving production processes.
Top Skills: AutocadControl Logix PlcsControls SystemsFanuc RoboticsHmisMaximoMS OfficeRockwellVariable Frequency Drives
Mid level
Automotive • Big Data • Information Technology • Robotics • Software • Transportation • Manufacturing
The Marketing Finance Analyst will design, execute, and oversee incentive programs, optimize financial performance, and manage projects to enhance strategic decisions.
Top Skills: ExcelPower BI

What you need to know about the Vancouver Tech Scene

Raincouver, Vancity, The Big Smoke — Vancouver is known by many names, and in recent years, it has gained a reputation as a growing hub for both tech and sustainability. Renowned for its natural beauty, the city has become a magnet for professionals eager to create environmental solutions, and with an emphasis on clean technology, renewable energy and environmental innovation, it's attracted companies across various industries, all working toward a shared goal: advancing clean technology.

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account