Software Engineer - Research Data Platform
Job Description
Key Skills Required
Master these to land this role
Want to know if you're a match for this job?
Bioptimus is building the first universal AI foundation model for biology to fuel breakthrough discoveries and accelerate innovation in biomedicine. With more than $75M in funding, Bioptimus is a fast-growing startup incorporated in October 2023 and headquartered in Paris. Backed by leading international venture capitalists, our world-class team of scientists and engineers is redefining the frontiers of AI and life sciences.
This is a remote role. Weâre headquartered in Paris, but the position can be performed remotely outside of Paris.
About the Role
We are a fast-moving, data-centric startup on a mission to bridge the gap between complex biological data and cutting-edge AI. As a Software Engineer in our research data platform team, you will help develop the backbone of our data architecture, designing and scaling the systems that power our AI models and user-facing tools, both internal and external.
We are looking for someone passionate about scalable, efficient, and highly structured data storage. In particular, we are looking for someone interested in designing systems that account for the complex structures inherent in biological data. You will build clean, maintainable systems that make massive biological datasets accessible, reliable, and actionable. If you love optimizing performance, improving schemas, and seeing your work directly empower a broad audience of stakeholdersâfrom scientists and product engineers to AI agentsâyou will fit right in.
This is a mid-level to senior individual contributor role. You will collaborate closely with our multidisciplinary team of researchers and engineers to drive software development and productization efforts.
What You Will Be Doing
As a Software Engineer for our research data platform, you will own the following responsibilities:
- Architect and build for performance: Design, implement, and maintain robust, scalable data schemas and storage solutions optimized for high-performance AI workloads.
- Optimize storage formats: Benchmark, profile, improve, and extend distributed storage using chunking, compression, parallelization, and custom solutions.
- Produce clean code: Develop and maintain high-quality, production-ready data systems, following clean code principles and engineering best practices.
- Design interfaces for people and AI agents: Build clean, typed, well-documented APIs that let researchers, engineers, and AI agents query and extend the platform programmatically.
- Collaborate and drive delivery: Work with researchers and product engineers to scope needs, align on priorities, and own projects end-to-end.
- Performance tuning: Monitor, profile, and optimize database queries, storage read and write paths, pipeline bottlenecks, and cloud infrastructure costs.
- Data governance and security: Collaborate with platform engineers to implement rigorous data validation, testing, versioning, and access control.
What You Will Bring
The successful candidate will have a team-first attitude, be independent, curious, and detail-oriented, thrive in a dynamic, fast-paced environment, and be fun to work with. Moreover, we value individuals with the following skills:
Technical and Professional Qualifications
- Python expertise: Deep, production-level knowledge of Python with a passion for clean, readable, and highly maintainable code.
- Backend frameworks: Strong hands-on experience with modern Python data tools and frameworks, such as Pydantic (data validation), SQLAlchemy (ORM), Alembic (database migrations), object storage abstractions, and FastAPI or similar frameworks.
- Structured databases: Expertise in relational database management systems (RDBMSs) such as PostgreSQL, including schema design, indexing strategies, and query optimization.
- Interfaces and agents: Experience designing API surfaces and exposure to protocols for programmatic and agent access such as the Model Context Protocol (MCP).
- User-centric mindset: A strong belief that data infrastructure is a product, combined with a commitment to keeping it usable and accessible to non-technical stakeholders.
How to Stand Out
Each of the following would be a valuable bonus, not a requirement:
- Biotech/life sciences affinity: Prior experience handling biological data formats (e.g., histology, transcriptomics, genomics, proteomics, or clinical trial data) or working in a biotech/health-tech environment.
- Start-up agility: A proven track record of thriving in fast-paced, ambiguous startup environments in roles requiring high autonomy and ownership.
- Array and storage formats: Experience with efficient distributed array storage (e.g., xarray, Zarr, TileDB, and TIFF) for both dense and sparse data, and comfort working close to library internals.
- Workflow orchestration: Experience with orchestration tools such as Dagster, Airflow, Prefect, or database-backed work queues.
- Frontend and visualization: Experience building or integrating with frontend and visualization tools that make data explorable.
- Proactive communicator: Ability to translate complex data architecture concepts into clear explanations for scientists, product managers, and engineers alike.
If your strengths lie in just one of these areas and you are passionate about biological data and scalable systems, we highly encourage you to apply!
How would you rate this job post?
See what other professionals think about this role.
Similar Opportunities
More Openings at Bioptimus
Explore Top Companies in this Space
SAIGroup
Enterprise AI Private Equity & Venture Studio / Predictive & Generative AI Vertical Software / Healthcare, Fintech & Retail Digital Transformation
The Public Interest Company
Healthcare FinTech & Subrogation Automation SaaS / InsurTech Third-Party Liability AI / Clinical Claims Auditing & Financial Recovery Infrastructure
UpSmith
Artificial Intelligence / Enterprise Software / Home Services / SaaS
Salvo Software
Enterprise Software / ERP Systems / Business Automation / IoT
Bioptimus
View Company ProfileBioptimus (operating at bioptimus.com) is a pioneering AI foundation model company engineered for transforming biology. Founded in 2024 by ex-Google DeepMind and Owkin scientists, Bioptimus is headquartered in Not specified, Bioptimus builds the first universal AI foundation model in biology to drive advancements in the field. Under the hood, Bioptimus trains foundation models natively across modalities and scales â learning the dynamics of human biology to predict what happens next. This allows researchers and scientists to gain deeper insights and more powerful performance than was possible before. Backed by $76M in funding, led by Cathay Innovation.
Safety First
- Never pay for a job application.
- Do not share sensitive bank info.
- Verify the client before starting work.




