Opens in a new tab

Blog

CDMO Evaluation Criteria: 12 Factors to Score Before You Sign

Biopharma manufacturing facility used to assess CDMO evaluation criteria, capabilities, and operational readiness.

Selecting the wrong CDMO causes catastrophic timeline slips, quality findings, tech-transfer pain, and commercial supply risks. Subjective preferences cannot justify critical outsourcing decisions. This scorecard-first framework outlines the CDMO Evaluation Criteria: 12 Factors to Score Before You Sign, equipping procurement, CMC, quality, and ops leaders to make defensible, audit-ready choices using phase-specific weighting and key RFP evidence.

Review our guide on how to choose a CDMO for broader context before diving into Factor #1.

1. Evaluate Site-Specific Compliance and Inspection History

CDMO compliance evaluation reviewing FEI numbers, inspection history, Form 483s, EIR summaries, and CAPA evidence.

A corporate brand means nothing if your assigned manufacturing facility holds open regulatory enforcement actions. Establish a pass/fail kill-gate for active enforcement before scoring any site.

  • What to Score (1 to 5): Inspection outcome trends, finding severity, and recurrence rates.
  • Evidence to Request: Facility FEI number, 3 to 5 year inspection history, Form 483s with responses, EIR summaries, and CAPA effectiveness evidence.
  • Red Flags: Active Warning Letters, repeat Quality Unit or data integrity findings, and evasive compliance responses.

2. Assess Quality Management System (QMS) Maturity

Move beyond basic compliance to score whether a CDMO’s QMS operates predictably under stress, reducing batch rejections and supply risks.

  • What to Score (1 to 5): Deviation rigor, batch record review cycle time, change control discipline, and audit trail reviews.
  • Evidence to Request: Recent quality metrics (right-first-time, repeat deviations), data integrity SOPs, and access controls.
  • Red Flags: “Paper GMP” responses, weak root-cause analysis, uncontrolled spreadsheets, and OOS trends.

Audit how facilities detect, investigate, and prevent recurrence, rather than how quickly they close CAPAs.

3. Validate Relevant Modality Experience and Scale

Confirm the CDMO has proven success with your specific modality, unit operations, and critical quality attributes (CQAs) at scale to avoid costly tech transfer resets.

  • What to Score (1 to 5): Direct experience with your modality (small molecule, biologic, viral vector), process complexity, and CQAs.
  • Evidence to Request: Sanitized case studies, equipment ranges, scale-up methodologies, and SME resumes.
  • Red Flags: “We can do anything” claims, missing references, or pushing standard processes without risk justification.

Review biologics-specific considerations for complex modalities.

4. Benchmark Process Development and Problem-Solving Capability

Differentiate CDMOs that actively solve process problems from those that only execute fixed recipes. Technical problem-solving drives first-pass success when issues inevitably arise.

  • What to Score (1 to 5): DoE capability, root-cause depth, impurity strategies, comparability/bridging support, and dev-to-manufacturing handoffs.
  • Evidence to Request: Sanitized deviation investigations, development reports, tech transfer playbooks, and SME availability during campaigns.
  • Red Flags: Limited development bandwidth, “manufacturing-only” posture, and slow responses when issues emerge.

Phase Note: Early-phase programs weight process development higher than mature commercial supply.

5. Audit Facility Infrastructure and Contamination Controls

Biopharma professional documenting quality checks during a CDMO manufacturing inspection.

Verify the facility can safely process your product without cross-contamination or capacity constraints.

  • What to Score (1 to 5): Line segregation, HVAC airflow, cleaning validation methodology, HPAPI containment, and biosafety controls.
  • Evidence to Request: Facility walk-through agenda, equipment specs and ranges, cleaning validation strategy, and multi-product scheduling models.
  • Red Flags: Shared equipment lacking adequate controls, unclear contamination limits, and inability to articulate worst-case scenarios.

Kill-Gate: Insufficient containment or physical segregation for high-risk or potent compounds.

6. Score Capacity Realism and Long-Term Scale-Up

Program timelines fail when slot commitments rely on wishful thinking instead of real operational bandwidth.

  • What to Score (1 to 5): Slot date realism, staffing stability, scale-up pathways, and Phase III or commercial capacity reservations.
  • Evidence to Request: Master schedules, changeover assumptions, capacity reservation terms, and site surge plans.
  • Red Flags: “Start next month” promises without a schedule, single-planner dependencies, or no path to scale.
  • Negotiation Hook: Treat capacity reservation terms as a core commercial scoring factor rather than a technical preference.

7. Score Proven Operational Performance Over Promises

Evaluate actual execution over sales claims to safeguard clinical timelines and commercial launch readiness.

  • What to Score (1 to 5): On-time in-full (OTIF) delivery, right-first-time batch success, investigation cycle times, and repeat deviation rates.
  • Evidence to Request: Sanitized KPI dashboards, explicit OTIF service-level definitions, and sample schedule recovery plans.
  • Red Flags: Missing metrics, midstream definition changes, and chronic “minor” deviations that consume sponsor bandwidth.

RFP Tip: Require bidders to use identical KPI definitions to prevent apples-to-oranges comparisons.

8. Assess Analytical Method Capabilities and Stability Infrastructure

Weak analytical capabilities stall product releases and regulatory submissions. Score whether analytics will accelerate your timeline or become a bottleneck.

  • What to Score (1 to 5): In-house testing capabilities, stability infrastructure, compendial alignment, and method transfer and validation experience.
  • Evidence to Request: Test menus, equipment lists, lead times, method transfer protocols, and OOS handling procedures.
  • Red Flags: Uncontrolled routine testing outsourcing, vague reference standard strategies, and delayed stability initiation.

Phase Note: Late-stage programs require detailed data packages and assay lifecycle management.

9. Evaluate Tech Transfer Systems and Joint Governance

Score the transfer system rather than a polished pitch deck. Successful transfer relies on co-owned execution, and system readiness strongly predicts first-batch success and validation outcomes.

  • What to Score (1 to 5): Tech transfer plan quality, joint steering committee (JSC) governance, acceptance criteria clarity, and closure reporting discipline.
  • Evidence to Request: Sample transfer plan, RACI matrix, method transfer protocol, engineering run strategy, and closure report template.
  • Red Flags: “Send us your batch record and we’ll run it” claims, vague responsibilities, and missing closure documentation.

10. Score Supply Chain Resilience and Disruption Risk

Supply chain analyst monitors global shipment routes on digital maps to help protect clinical trial supply chains from geopolitical risk and disruptions.

Single points of failure in materials, shipping, or QA release can halt a clinical program. Evaluate how proactively a CDMO eliminates disruption risks.

  • What to Score (1 to 5): Dual-sourcing depth, safety stock management, supplier qualification, and cold-chain logistics.
  • Evidence to Request: Critical material lists, supplier SOPs, lead-time assumptions, and temperature excursion protocols.
  • Red Flags: Opacity around sub-tier suppliers or placing all risk onto the sponsor without joint mitigation.
  • Commercial Tie-In: Define material shortage liability and cap pass-through markups in the agreement.

11. Evaluate Commercial Transparency and Contractual Protections

A low headline price often masks costly change orders and scope creep. Evaluate contracts on financial clarity and risk allocation.

  • What to Score (1 to 5): Pricing transparency (FTE vs. batch), change-order governance, batch failure liability, material markups, and MOQs.
  • Evidence to Request: Rate card logic, pass-through policies, tech transfer fees, data ownership, and exit rights.
  • Red Flags: Take-it-or-leave-it MSAs, IP or exit refusal, undefined markups, and liability caps misaligned with fault.

Practical Rule: Score contracts on structural fairness, not initial price.

12. Assess Collaboration, Governance, and Responsiveness

Collaboration functions as a measurable operational control, not a soft skill. Weak governance and slow communication reliably predict delays during tech transfer and clinical supply execution.

  • What to Score (1 to 5): PM structure, escalation pathways, meeting cadence, documentation discipline, and decision turnaround times.
  • Evidence to Request: Sample status reports, risk register templates, escalation SOPs, and team continuity commitments.
  • Red Flags: Unclear ownership, rotating personnel, slow RFP responses, and a defensive posture during audits.

Post-selection, apply the same rigor to ongoing governance — joint KPIs, escalation paths, and defined roles keep the partnership on track after the contract is signed.

How to Operationalize a Repeatable CDMO Selection Workflow

Transform evaluation criteria into a repeatable procurement process that produces an audit-ready decision trail and a defensible shortlist.

Step 1: Establish Pass-Fail Kill Gates Before Scoring

Filter out non-viable suppliers during the preliminary RFI stage. Define non-negotiable disqualification criteria, such as active FDA Warning Letters, inadequate physical containment for potent compounds, inability to meet target slot dates, or missing core analytical capabilities.

Output: A concise disqualification checklist used to eliminate unsuitable candidates before detailed scoring.

Step 2: Define a Standardized 1 to 5 Scoring Legend

Establish clear evaluation definitions to maintain consistency across internal reviewers:

  1. High risk with little to no supporting evidence.
  2. Acceptable performance with minor gaps requiring remediation.
  3. Strong documentation, reliable capabilities, and repeatable performance.

Require evaluators to cite concrete evidence for each score, referencing specific document numbers, operational metrics, or audit observations.

Step 3: Apply Phase-Appropriate Scorecard Weighting

Tailor evaluation weightings to align with your program’s lifecycle stage:

  • Baseline Weighting: Quality and Compliance (25%), Technical Fit (30%), Capacity, Timing, and Track Record (20%), Commercial Terms (15%), and Strategic Alignment (10%).
  • Early Clinical Phase: Upweight development speed, technical troubleshooting, and operational flexibility.
  • Late Phase and Commercial: Upweight regulatory inspection history, formal capacity reservations, supply chain resilience, and contractual protections.

Step 4: Run the RFP as a Structured Data Collection Exercise

Standardize response templates to compare commercial proposals and technical data side by side. Flag non-responsiveness or evasive answers as scored risk signals.

Step 5: Shortlist Bidders and Verify Through Focused Audits

Use the lowest-scoring domains from your evaluation scorecard to focus on-site audits rather than conducting generic facility tours. Establishing a supported execution pathway means combining structured diligence with clear governance from day one — contact our team to build one for your program.

Biopharma professionals reviewing CDMO information together on a laptop during partner evaluation.

Build Your Custom CDMO Selection Process

Our team can help you build, weight, and operationalize a defensible CDMO scorecard tailored to your next manufacturing campaign.

Contact Our Team

Frequently Asked Questions About CDMO Evaluation Criteria

What weights should we use in a CDMO scorecard?

A baseline CDMO scorecard typically allocates 25% to Quality and Compliance, 30% to Technical Fit, 20% to Capacity and Track Record, 15% to Commercial Terms, and 10% to Strategic Alignment. Sponsors should adjust these weights based on project phase and run a sensitivity analysis to see if shifting Quality or Technical Fit weights by 5% to 10% changes the top vendor.

How do we evaluate a CDMO’s FDA Form 483 or Warning Letter history without overreacting?

Evaluate compliance history by confirming site-specific records, requesting 3 to 5 years of inspection artifacts, and scoring the quality of vendor responses rather than just the initial findings. Use active enforcement actions as immediate kill-gates, but focus on response rigor, root-cause depth, and whether issues recur when evaluating standard Form 483 observations.

What are common CDMO RFP red flags?

Key red flags during the RFP process include opaque pricing models, refusal to share quality metrics, evasive communication, and overpromised schedules lacking realistic slot commitments. Inflexible master services agreements (MSAs) or a reluctance to negotiate basic joint governance terms also signal potential operational friction during technology transfer and campaign execution.

How should evaluation criteria change from early clinical phases to commercial supply?

Early-phase evaluations prioritize speed, technical troubleshooting, process development capability, and flexible slotting. Commercial-stage evaluations shift heavily toward proven regulatory compliance histories, continuous process verification maturity, formal capacity reservations, supply chain resilience, and enforceable contractual protections.

When should we bring in outside support for CDMO selection?

Engage external expertise when internal teams lack specialized technical bandwidth, face inconsistent scoring across cross-functional stakeholders, or manage high-stakes, multi-site selections. Strategic advisors help build defensible decision frameworks and prevent costly oversight errors. Explore our resources on how to choose a CDMO and managing CDMO relationships, or contact Syner-G for expert selection and oversight support.

Score every CDMO candidate against these 12 factors before you sign, and you replace subjective vendor selection with a defensible, audit-ready decision trail. Weight the criteria to your program phase, verify claims with real evidence, and treat any kill-gate red flag as a hard stop.

Book a Free Consultation

Share

Related Resources

All Resources