Company Description IntegriChain is the data and application backbone for market access departments of Life Sciences manufacturers. We deliver the data, the applications, and the business process infrastructure for patient access and therapy commercialization. More than 250 manufacturers rely on our ICyte Platform to orchestrate their commercial and government payer contracting, patient services, and distribution channels. ICyte is the first and only platform that unites the financial, operational, and commercial data sets required to support therapy access in the era of specialty and precision medicine. With ICyte, Life Sciences innovators can digitalize their market access operations, freeing up resources to focus on more data-driven decision support. With ICyte, Life Sciences innovators are digitalizing labor-intensive processes - freeing up their best talent to identify and resolve coverage and availability hurdles and to manage pricing and forecasting complexity. We are headquartered in Philadelphia, PA (USA), with offices in: Ambler, PA (USA); Pune, India; and Medelln, Colombia. For more information, visit , or follow us on Twitter @OpenKyber and LinkedIn. This role offers flexibility, but candidates must reside in Pennsylvania, New Jersey, or New York and be within a reasonable travel distance of our Philadelphia office, as regular in-person collaboration is required.
Job Description Mission Join the Data Science team as an AI Data Engineer responsible for building the data foundations that make enterprise AI products accurate, explainable, and scalable. This role will design and implement Snowflake and dbt pipelines from raw source data to curated gold-layer datasets, create semantic models that LLM tools can use reliably, and partner with data science, product, and engineering teams to convert data dictionaries and business definitions into AI-ready data products. The ideal candidate is a strong data engineer with deep Snowflake/dbt experience and a practical understanding of how semantic layers, ER relationships, denormalized models, and metadata quality influence LLM and agent performance.
6+ years of experience in data engineering, analytics engineering, database engineering, or data platform development in production environments. Strong hands-on experience with Snowflake, including SQL development, performance tuning, security-aware design, cost optimization, and large-volume processing. Strong hands-on experience with dbt or comparable ELT tooling, including models, tests, documentation, lineage, and environment promotion. Experience building raw-to-curated-to-gold data pipelines and business-ready datasets. Strong SQL and Snowflake development skills, including complex transformations, views, stored procedures/Snowflake Scripting, and query optimization. Experience creating semantic layers, semantic models, metrics, dimensions, relationships, and curated analytical views. Good understanding of ER modeling, dimensional modeling, denormalized consumption models, and data grain management. Experience translating data dictionaries and business definitions into physical models, dbt models, and semantic-layer definitions. Understanding of Snowflake Cortex capabilities such as Cortex Analyst, Cortex Complete, and semantic-model-driven natural-language querying. Ability to partner with data science, product, engineering, and business teams to deliver AI-ready data products.
Preferred Experience
Experience in life sciences, healthcare, pharma commercialization, MDM, patient data, channel data, or commercial data platforms. Experience with Snowflake semantic views, Cortex Analyst, Cortex Search, or other AI/LLM data platform capabilities. Experience with data quality frameworks, metadata management, data observability, and lineage tooling. Experience with orchestration tools such as dbt Cloud jobs, Airflow, Dagster, cloud-native schedulers, or similar platforms. Experience with Python for data automation, metadata processing, testing, or API integrations. Experience designing governed data products for BI, AI/ML, natural-language analytics, or agentic applications. Snowflake SnowPro, dbt certification, or equivalent data engineering credentials.
Additional Information
What does OpenKyber have to offer? Mission driven: Work with the purpose of helping to improve patients' lives! Excellent and affordable medical benefits + non-medical perks including Student Loan Reimbursement, Flexible Paid Time Off and Paid Parental Leave 401(k) Plan with a Company Match to prepare for your future Robust Learning & Development opportunities including over 700+ development courses free to all employees #LI-ZG1 OpenKyber is committed to equal treatment and opportunity in all aspects of recruitment, selection, and employment without regard to race, color, religion, national origin, ethnicity, age, sex, marital status, physical or mental disability, gender identity, sexual orientation, veteran or military status, or any other category protected under the law. OpenKyber is an equal opportunity employer; committed to creating a community of inclusion, and an environment free from discrimination, harassment, and retaliation. Our policy on visa sponsorship for US based positions: Applicants for employment in the US must have valid work authorization that does not now and/or will not in the future require sponsorship of a visa for employment authorization in the US by OpenKyber.
For applications and inquiries, contact:[email protected]
By continuing you agree to our Terms & Privacy Policy.