Every Horizon Europe applicant eventually holds the same two-page document: the Evaluation Summary Report, or ESR, with three scores and three blocks of comments. Most read it once, in the first hour after the result letter, and never decode it. That is a waste, because the ESR is written to a published standard by experts who were briefed with a published slide deck, and once you know the rules the report tells you precisely why the proposal landed where it did and what a resubmission has to change. This page sets out those rules from the Commission’s own evaluator briefing and then reads two large samples of real ESRs against them.
The evaluation in four steps
After the admissibility and eligibility check, a proposal passes through three expert stages. In the individual evaluation, each expert, usually three or more, reads the proposal alone and completes an Individual Evaluation Report with a comment and a score for each criterion. The briefing tells them to evaluate the proposal as submitted, not on its potential, to reflect any shortcoming beyond minor ones and clerical errors in a lower score, to make scores match comments, and to explain shortcomings without making recommendations. In the consensus step the same experts meet, moderated by Commission or agency staff, and agree comments before agreeing scores; for full proposals they are told not to converge immediately on the average, while for first-stage proposals the average is the starting point. A rapporteur drafts the consensus report. In the panel review a wider panel checks consistency across proposals, resolves minority views, cross-reads proposals with equal scores, endorses final scores and comments, and recommends a ranked list. The briefing is explicit that the consensus report often remains unchanged at panel stage, so in most cases the ESR you receive is the consensus report verbatim.
The scale, in the words experts are given
Scores are awarded per criterion, not per aspect within a criterion, in steps of 0.5 from 0 to 5, for a maximum of 15. The briefing asks experts to use the whole range and defines each integer as follows.
| Score | Descriptor | Definition given to evaluators |
|---|---|---|
| 0 | Fails | The proposal fails to address the criterion or cannot be assessed due to missing or incomplete information. |
| 1 | Poor | The criterion is inadequately addressed, or there are serious inherent weaknesses. |
| 2 | Fair | The proposal broadly addresses the criterion, but there are significant weaknesses. |
| 3 | Good | The proposal addresses the criterion well, but a number of shortcomings are present. |
| 4 | Very good | The proposal addresses the criterion very well, but a small number of shortcomings are present. |
| 5 | Excellent | The proposal successfully addresses all relevant aspects of the criterion. Any shortcomings are minor. |
Minor shortcoming, shortcoming and significant weakness are defined terms in the evaluator briefing, not loose description, and they are the key to reading a comment. Each one tells you how far the score moved and whether the proposal is still fundable.
| Term | Definition given to evaluators | Effect on the score | How it reads in an ESR |
|---|---|---|---|
| Minor shortcoming | Relates only to a marginal aspect, or can easily be rectified. | None | “The proposal would benefit from...”, alongside an otherwise positive paragraph |
| Shortcoming | Relates to an important aspect, but does not make the proposal inappropriate for funding. | Lowers the score; roughly 3.5 to 4 | “not sufficiently detailed”, “partially addressed” |
| Significant weakness | The criterion is addressed in a limited or not sufficiently effective way. A large number of individual shortcomings can together amount to one. | Pushes the score below threshold; 2.5 or below | “not addressed”, “not convincing” |
The practical consequence is that a single significant weakness on any criterion is fatal on its own, while a page of minor shortcomings is not. Count the defect words in your ESR before counting the comments.
Thresholds, weighting and the 3/3/3 trap
The default threshold is 3 for each criterion and 10 for the total, unless the work programme says otherwise, and a proposal must pass both to be considered for funding. The arithmetic has a consequence that surprises many first-time applicants: three scores of exactly 3 pass every individual threshold, total 9, and are rejected. “Good” everywhere is not fundable. In practice a fundable full proposal needs at least one 4 and no score below 3, and in competitive topics the funding line sits well above 10.
| Submission type | Criteria scored | Threshold per criterion | Overall threshold | Weighting |
|---|---|---|---|---|
| Single-stage or second-stage full proposal (RIA, IA, CSA) | Excellence, Impact, Implementation | 3 / 5 | 10 / 15 | Normally none; where a type of action weights a criterion it counts only for ranking |
| First stage of a two-stage call | Excellence, Impact only | 4 / 5 | Set so that the requested budget of passing proposals is a multiple of the call budget | None |
| MSCA Doctoral Networks from the 2026 call | Excellence, Impact, Implementation (percentage scores) | Minimum in each criterion introduced in 2026 | Call-specific | MSCA-specific weights |
Passing is not being funded: ranking and ties
Once thresholds are met, proposals are ranked by total score and funded in order until the topic budget is exhausted. Those above threshold but outside the budget are placed on a reserve list, and in calls that provide it, receive a Seal of Excellence that some national and regional funders honour. Because many proposals cluster at 13 to 14.5, the tie-break order matters more than applicants expect. For each group of proposals with the same total, the panel applies the following sequence.
- Proposals addressing aspects of the call not otherwise covered by higher-ranked proposals are prioritised.
- Within that group, priority goes to the higher Excellence score, then Impact. For Innovation Actions the order is Impact first, then Excellence.
- If still tied, the gender balance among the personnel named as primarily responsible in the researchers table.
- Then geographical diversity: the number of Member States or Associated Countries in the proposal not otherwise funded higher up the list, and if equal, budget.
- Finally, other factors related to the call or to Horizon Europe generally, such as portfolio synergies or SME involvement.
Two practical consequences follow. For a RIA, an extra half point on Excellence is worth more than the same half point on Implementation, because it breaks ties. And the researchers table in Part A is not decorative: it can decide a tie at step 3.
What the ESR contains, and what it is not allowed to contain
The ESR arrives with the evaluation result letter through the Funding & Tenders Portal; the programme’s service standard is to inform applicants within five months of the deadline and to sign grants within eight. The report has a main part with comments and a score for each of the three criteria, plus any remarks on scope, on activities excluded from funding, and on operational capacity. The briefing sets a quality standard that experts are reminded applicants will read and can use in an evaluation review. The report must reflect strengths and weaknesses fairly and give reasons for the scores; it must address all aspects listed under each criterion and only those; it must not contain comparative statements about other proposals, because each proposal is judged on its own merit; it must not penalise the same weakness under two criteria; and it must not contain recommendations or suggestions for improvement. Comments must be precise, verified, and consistent as a whole with the meaning of the score awarded.
Three further rules are useful to applicants. For second-stage proposals, inconsistencies between the stage-1 and stage-2 ESR must be specifically justified. For resubmissions declared in the form and made within two years under comparable conditions, comments and scores that differ significantly from the previous ESR must also be justified. And a proposal in a blind first stage that directly identifies an applicant in Part B or the abstract is declared inadmissible before any of this happens; the rules and a checklist are in the two-stage and blind calls guide.
What real ESRs say: two analyses
Two National Contact Point teams have read large samples of ESRs and coded the comments. The Horizon Academy project, led by the Slovak NCP centre CVTI SR, analysed 130 reports from December 2021 to March 2024, taking the five highest- and five lowest-scored ESRs per call across every Pillar II cluster and the Widening component, for RIA, IA and CSA actions. The French Health NCP analysed 189 reports from the 2021 Health cluster RIA topics with a total score of 11 or more, out of 501 proposals submitted, 207 of which, or 41 percent, fell below the overall threshold. The recurring comments line up closely with the aspects in the template, which is the point: experts are answering the template’s questions, and the comments record which ones went unanswered.
| Criterion | Aspect | Positive comment pattern (4-5) | Negative comment pattern (below 3.5) |
|---|---|---|---|
| Excellence | Objectives | Clear, measurable, SMART; KPIs quantified with method of calculation; explicitly linked to the topic's expected outcomes | Objectives not linked to call outcomes; no KPIs or KPIs without explanation; state of the art not described; score 'drops quickly' if judged not beyond the state of the art |
| Excellence | Methodology | Concepts, assumptions and hypotheses articulated and justified; TRL start and end credible; interdisciplinarity explained | Generic or unconvincing methodology; missing justification; SSH integration absent where the topic asked for it; AI role, robustness and training data not described |
| Excellence | Open science and gender | Addressed adequately as an integral part of the methodology; FAIR data management described | Treated as boilerplate or omitted; most often a positive-comment area, but a silent proposal loses the point |
| Impact | Pathway to impact | Credible, detailed, quantified; barriers identified with mitigation; the single most commented aspect in both samples | 'Not sufficiently described', 'not ambitious', 'not convincing', 'not quantified'; scalability not considered; barriers not identified or mitigation generic |
| Impact | Dissemination, exploitation, communication | Measures tailored to each target group (patients, professionals, policy-makers); IP strategy detailed; go-to-market plausible for IAs | Generic plan; target groups not segmented; IP management missing; commercialisation path absent in an IA |
| Implementation | Work plan | Detailed, logical WP sequence with coherent flow; Gantt and PERT present; milestones usable for monitoring | WPs not detailed, inconsistencies or overlap between WPs, timeline inappropriate; deliverables and milestones not matched to tasks; 40 percent of negative Implementation comments in one destination |
| Implementation | Effort and budget | Budget and person-months transparent and proportionate; all partners meaningfully involved; subcontracting justified | Tasks over- or under-estimated; unbalanced budget across partners or WPs; personnel-heavy WPs; subcontracting not justified; PM tables inconsistent with Part A |
| Implementation | Risks | Complete list, risks well identified, credible and specific contingency plan | Risk analysis under-elaborated or under-estimated; AI-specific and clinical-trial risks (recruitment, regulatory delay) not covered; contingency generic |
| Implementation | Consortium | Adequate capacity and resources; history of collaboration; relevant SME participation; scientific advisory board | Missing SME or industrial partners; roles of some participants undefined; a domain of expertise missing; effort distribution unbalanced without explanation |
Two findings from the samples are worth stating on their own. The Horizon Academy authors note that where a proposal is written understandably, makes a good impression and covers almost all requested aspects, experts are tolerant of minor shortcomings and will give the top score; the losses come from aspects that were simply absent. The French team found little difference between funded and unfunded proposals in thenumber of negative comments on the most commented aspects; what differed was whether those comments used the vocabulary of shortcomings or of significant weaknesses. Read your ESR for that vocabulary, not for the count.
Reading your own ESR: the resubmission decision
Put the three scores next to the thresholds and the funding line for your topic, which the result letter or the NCP can tell you, and classify the outcome. The bands below are a working heuristic built on the descriptor definitions and the two ESR samples, not an official rule.
| Outcome | What the ESR is telling you | Sensible response |
|---|---|---|
| Any criterion below 3 | A significant weakness: the criterion was addressed in a limited or ineffective way. The comment names the missing aspect. | Rebuild that section from the template question outward. Do not resubmit with cosmetic edits; the same experts may see it and must justify any large change in score. |
| All criteria at or above 3 but total below 10 (for example 3/3/3 or 3.5/3/3) | No single failure, but a number of shortcomings everywhere. The proposal is 'good' and unfundable. | Pick the two criteria with the most concrete comments and lift each by a full point: add KPIs, quantify the pathway, fix the WP tables. Consider a different topic if the scope comment is lukewarm. |
| Total 10 to 12, above thresholds, not funded | Eligible but well below the funding line. Comments are mostly shortcomings, often in Impact. | Resubmit to the next matching topic after addressing every named shortcoming, and declare the resubmission so the new panel must justify any divergence. |
| Total 12.5 to 14, reserve list or just under the line | Fundable quality, lost on ranking. Check the tie-break: Excellence for RIA, Impact for IA. | Strengthen the tie-break criterion by half a point, keep everything the experts praised, and resubmit. If a Seal of Excellence was issued, approach national funders that honour it. |
| Total 14 or above, not funded | Budget exhausted above you, or a scope decision under step 1 of the tie-break. | Resubmit essentially unchanged to the next call, updating only dates, state of the art and any comments flagged as shortcomings. |
Whichever band you are in, the rewrite should start from the ESR comment, then the corresponding template question in the Part B template walkthrough, and only then the old text. If the comments point at implementation tables, the work package and Gantt worked example shows what a fully specified plan looks like. For the broader strategy of when a resubmission is worth the effort, see The Resubmission Renaissance.
Frequently asked questions
- What is the passing score for a Horizon Europe proposal?
- By default each of the three criteria must score at least 3 out of 5 and the total must reach at least 10 out of 15. Both conditions apply. Three scores of exactly 3 pass every individual threshold but total 9, which fails the overall threshold. A work programme topic can set different thresholds, so always check the call conditions.
- What is an Evaluation Summary Report (ESR)?
- The ESR is the document you receive with the evaluation result letter. It contains, for each criterion, the agreed comments on strengths and weaknesses and the score, plus the total score and, where relevant, notes on admissibility, scope or activities excluded from funding. In most cases it is identical to the consensus report drafted by the rapporteur of the expert group.
- Does a score above the thresholds guarantee funding?
- No. Passing the thresholds only makes a proposal eligible for funding; it is then ranked against the others in the call and funded in priority order until the budget runs out. Proposals above the thresholds but outside the budget go on a reserve list or receive a Seal of Excellence where the call provides one.
- How are proposals with the same total score ranked?
- The panel first prioritises proposals that address aspects of the call not covered by higher-ranked proposals. Within that group it ranks by the Excellence score, then Impact; for Innovation Actions the order is Impact first, then Excellence. If still tied it considers gender balance among the named researchers, then geographical diversity, then other call objectives such as SME involvement.
- Are the criteria weighted?
- Scores are normally not weighted. Where a type of action uses weighting, for example a higher weight on Impact for Innovation Actions, it is applied only to build the ranking, never to decide whether a proposal passed the thresholds.
- How long after the deadline does the ESR arrive?
- Horizon Europe commits to informing applicants of the evaluation outcome within five months of the call deadline and to signing grant agreements within eight months. The ESR is attached to the evaluation result letter sent through the Funding and Tenders Portal.
- Can I challenge my ESR?
- You can request an evaluation review within the deadline stated in the result letter, but only on procedural grounds: for example, that the evaluation was not carried out according to the published process or that an admissibility check was wrong. The review committee does not re-score the scientific merit of the proposal.
- If I resubmit, will the new evaluators see my old ESR?
- If you declare the resubmission in the application form and it is within two years, the briefing tells experts that comments and scores differing significantly from the previous ESR must be specifically justified when the resubmitted proposal was produced under comparable conditions. That is a reason to declare it, provided you have actually fixed the weaknesses.
Sources
- European Commission, Standard briefing slides for Horizon Europe expert evaluators — PDF, version 14.0, 15 April 2026; score descriptors, thresholds, consensus and panel rules, quality standard for ESRs, tie-break order, blind evaluation annex
- European Commission, Horizon Europe Work Programme 2026-2027, General Annexes — Annex D award criteria, scores, thresholds and weighting; Annex F evaluation procedure
- European Commission, Standard Application Form (HE RIA and IA), Part B — PDF, version 5.1, 22 January 2026; the aspects listed under each criterion
- Horizon Academy (NCP4HE), Report on ESR Analysis and Methods, Deliverable 3.1 — K. Papanova, CVTI SR, September 2024; 130 ESRs from December 2021 to March 2024 across RIA, IA and CSA
- PCN Santé (French Health NCP, MESRI), Analyses des ESR, appels 2021 — PDF, June 2022; 189 ESRs from the 2021 Health cluster RIA topics with total score of 11 or more
- Horizon Europe NCP Portal, Analysis of Evaluation Summary Reports (repository entry) — Landing page for the Horizon Academy analysis
- European Commission, Funding & Tenders Portal Online Manual: evaluation review procedure — How to request an evaluation review
Founder & CEO, Proposia.ai
PhD researcher and Associate Professor in Computer Science, working at the intersection of algorithm design, applied mathematics, and machine learning. With Proposia.ai, I aim to transform research ideas into scalable AI solutions that support innovation and discovery.