
Small Molecule Data Analyst
- On-site
- Suzhou, Jiangsu Sheng, China
- Product & Engineering: Life Sciences
Job description
We are hiring a Small Molecule Data Analyst who can bridge pharmaceutical R&D expertise, data analytics, and AI-driven data processing to deliver high-quality biomedical and chemical data solutions. This role translates complex client and business needs into actionable data requirements, collaborates with engineering, data quality, and AI teams to improve data pipelines and governance, and leads cross-functional data projects. The ideal candidate combines a strong chemistry or drug discovery background with hands-on experience in biomedical databases, SQL/Python, and Large Language Models, applying technologies such as prompt engineering and AI agents to extract, normalize, and mine insights from patents, scientific literature, and other unstructured data.
Job requirements
What You’ll Be Doing:
Pharmaceutical Data Analysis & Management: Familiar with the R&D process for chemical drugs. Able to accurately translate pharmaceutical clients' business needs into data product requirements.
Cross-functional Collaboration & Data Governance: Work closely with data engineering, data quality, and AI teams to ensure the efficient and stable operation of data mining and processing pipelines, guaranteeing the accuracy and usability of output data.
Data Quality System Development: Build and continuously refine the biomedical data quality evaluation system. Leverage customer feedback loops, data-driven analytical methods, and business rule optimization to consistently improve overall data quality.
Project Coordination & Management: Act as the primary interface for data operations. Coordinate cross-functional data projects and effectively manage project timelines and deliverables.
AI-Driven Data Processing: Proficient in Prompt Engineering with the ability to design and optimize prompts. Lead the application of Large Language Models (LLMs) to perform intelligent information extraction, normalization, and knowledge mining from unstructured content (e.g., patents, academic literature). Explore and build data processing Agents tailored for the biomedical field, such as drug R&D and patent strategy (patent layout).
About You:
Educational Background: Master’s degree or above in Synthetic Chemistry, Organic Chemistry, Computational Chemistry, Drug Design, AIDD (AI in Drug Discovery), or other chemistry-related fields.
Industry Experience: Candidates with data-related work experience in chemical drug R&D, clinical research, pharmaceutical analysis, or database construction at large domestic or global pharmaceutical companies are preferred.
Professional Skills: Proficiency in using biomedical/chemical databases such as SciFinder, Reaxys, STN, or GOSTAR is highly preferred.
Technical Skills: Ability to perform basic data processing using SQL, Python, etc. Experience in database design and practical implementation is a plus.
AI Technologies: Proficient in mainstream Large Language Models (LLMs) such as DeepSeek, GPT, Claude, and Gemini. Rich experience in AI applications and algorithms is preferred.
Comprehensive Abilities: Excellent logical thinking, business acumen, and fast-learning capabilities. Able to accurately translate complex business needs into data solutions. Outstanding cross-departmental communication, coordination, and project execution skills.
Language Skills: English as a working language (proficient in listening, speaking, reading, and writing). Excellent cross-cultural communication skills.
or
All done!
Your application has been successfully submitted!
You've already applied for this job
We appreciate your interest in this position. Unfortunately, you have already applied for this job.

