1,000+ Opportunities
Find the right grant
Search federal, foundation, and corporate grants with AI — or browse by agency, topic, and state.
RFP closed February 7, 2026; site explicitly states no longer accepting applications.
AI for Forecasting and Sound Reasoning is a grant from Open Philanthropy that funded projects using artificial intelligence to improve human reasoning and forecasting accuracy. The program supported tools and research for fact-checking, argument analysis, and decision-making under uncertainty, reflecting Open Philanthropy's interest in global catastrophic risk reduction and improving institutional reasoning.
This request for proposals has closed. Open Philanthropy is a major philanthropic funder focused on evidence-based giving across biosecurity, AI safety, global health, and other high-impact cause areas. Interested applicants may monitor Open Philanthropy's website for future related funding opportunities.
Get alerted about grants like this
Get emailed when new opportunities from “Coefficient Giving (formerly Open Philanthropy)” or related funders appear. Free, weekly, unsubscribe anytime.
Or search similar grants →Extracted from the official opportunity page/RFP to help you evaluate fit faster.
[Closed] Request for Proposals: AI for Forecasting and Sound Reasoning | Coefficient Giving if (e. matches && this. menuOpen) { // Use RAF for performant scroll updates window.
addEventListener('scroll', () => { window. requestAnimationFrame(() => { document. documentElement.
style. setProperty('--scroll-y', `${window. scrollY}px`); Biosecurity & Pandemic Preparedness Effective Giving & Careers Global Catastrophic Risks Opportunities Global Health & Wellbeing Opportunities Lead Exposure Action Fund Navigating Transformative AI Science and Global Health R&D Coefficient is a philanthropic funder and advisor.
*+*]:border-t-1 [&>*+*]:border-white-200 border-white-200 border-b-1"> Biosecurity & Pandemic Preparedness Effective Giving & Careers Global Catastrophic Risks Opportunities Global Health & Wellbeing Opportunities Lead Exposure Action Fund Navigating Transformative AI Science and Global Health R&D *+*]:mt-4 flex-1/2 self-start"> if (item. classList. contains('active')) { item.
style. anchorName = '--hovered'; const { width } = item. getBoundingClientRect(); this.
styles['--width'] = `${width}px`; item. style. anchorName = ''; const nav = $el.
querySelector('. subnav'); const items = [... nav.
querySelectorAll('. subnav__item')]; item. style.
anchorName = '--hovered'; item. style. anchorName = ''; const { width } = e.
target. getBoundingClientRect(); this.
styles['--width'] = `${width}px`; class="w-full hidden lg:block text-center relative"> RFP on AI for Forecasting RFP on AI for Forecasting [Closed] Request for Proposals: AI for Forecasting and Sound Reasoning *+*]:mt-2 list-none [&>li]:marker:content-[''] marker:content-['']"> Submitting your application What makes for a strong submission *+*]:mt-2 list-none [&>li]:marker:content-[''] marker:content-['']"> Further detail on our areas of interest *+*]:mt-2 list-none [&>li]:marker:content-[''] marker:content-['']"> Examples of projects we’re interested in funding *+*]:mt-2 list-none [&>li]:marker:content-[''] marker:content-['']"> Connecting Forecasting to Decision Processes Evaluations and benchmarks *+*]:mt-2 list-none [&>li]:marker:content-[''] marker:content-['']"> What kinds of progress are you expecting from AI forecasters?
What do you mean by high-stakes decision-making? Should I apply if I have never worked on GCRs, and don’t intend to? Can I submit more than one application?
Is there a limit on indirect project costs? What if I have other questions? — Submitting your application What makes for a strong submission Further detail on our areas of interest Examples of projects we’re interested in funding — Connecting Forecasting to Decision Processes — Evaluations and benchmarks — What kinds of progress are you expecting from AI forecasters?
— What do you mean by high-stakes decision-making? — Should I apply if I have never worked on GCRs, and don’t intend to? — Can I submit more than one application?
— Is there a limit on indirect project costs? — What if I have other questions? *]:lg:col-start-1!
[&>. wysiwyg:first-child_h2:first-child]:border-t-0 [&>. wysiwyg:first-child_h2:first-child]:mt-0"> Update (February 7, 2026): We are no longer accepting applications for this RFP.
We believe that modern AI has the potential to improve our ability to reason in a structured and quantified way, and thereby improve human decision-making. To help realize this potential, we are launching a request for proposals (RFP) for two related areas of interest: AI for forecasting: We are looking for proposals for AI models that help to make forecasts more accurate or more relevant.
We are primarily interested in probabilistic, judgmental forecasting, i.e. quantitative forecasts that cannot be based fully on large sets of structured data.
Aside from models that directly produce forecasts, ideally approaching or exceeding human performance on forecasting tasks, we’re also looking to fund work on models that perform one or more of the subtasks involved in using forecasts for decision-making, such as explaining the reasoning behind forecasts or building forecasting models. See here for a more detailed discussion of what we’re looking for.
AI for sound reasoning: Modern AI models are being adopted at a rapid pace throughout society, including for high-stakes decisions in law, academia and policy. We expect this trend to continue over the coming years, and possibly accelerate. It seems crucial to us that models that are used for highly consequential decisions are generally truth-oriented, and support such tendencies among their users.
We see two main paths to this goal that we’re interested in funding: Research into understanding when models do and do not support sound reasoning, including evaluations of models with respect to principles of sound reasoning like truthfulness, meta-reasoning, or consistency.
Developing tools that directly help with specific tasks that are disproportionately helpful for clear reasoning, like fact-checkers, fact tracers, arbitrators, or argument analyzers. See here for a more detailed discussion of what we’re looking for. We are primarily interested in the potential of AI models to improve high-stakes decision making, especially in the context of global catastrophic risks (GCRs) .
However, we expect that most projects we fund will also apply to decision-making with lower stakes, with only a minority of projects directly focusing on applications to GCRs. See our Q&A for more detail. Submissions will be evaluated by Coefficient Giving’s Forecasting team .
We plan to make grants totaling around $8-10 million for proposals submitted in this period. We expect the typical successful proposal will receive funding in the range of $100,000 to $1 million for a project duration between 6 months and 2 years.
While AI for forecasting is a subcategory of AI for sound reasoning in principle, we are more confident about our vision for the former, and currently expect to end up with more projects in this area. However, this assessment will be subject to continued reevaluation, and may very well be reversed in the future.
We’ve seriously considered the risk that supporting projects to improve forecasting and epistemic tools might inadvertently accelerate AI capabilities — e.g. by generating useful training data or by incentivizing benchmark-climbing.
However, given the relatively modest scale of funding, the focus on downstream tasks rather than fundamental model improvements, and the nature of the work (e.g. helping models extract, structure, and express knowledge more accurately), we believe the net effect is more likely to be positive.
In particular, these tools seem well-suited to improving model transparency, calibration, and alignment with truth — qualities we consider central to mitigating catastrophic risk. We’re also taking care to define the scope of this RFP so it avoids incentivizing the development of capabilities with agentic or deceptive potential, instead favoring tools that increase understanding and support good decision-making.
Submitting your application Update (February 7, 2026): We are no longer accepting proposals. If we re-open it, we will update this page with application materials. We will aim to respond to you within 6-8 weeks of receiving your proposal.
We may need additional time to respond in some cases — for example, if we decide to consult with external advisors, or if application volume is unexpectedly high. We do not plan to provide feedback for rejected proposals in most instances; we want to focus on evaluating the most promising proposals and responding quickly to all of our applicants.
If we move forward with your application, we will introduce you to the grant investigator who will evaluate your proposal in depth. They will then start a conversation with you, typically requesting further information as a first step. At this stage, you’ll have the opportunity to clarify and evolve your proposal in dialogue with the grant investigator and to develop a finalized budget.
See this page for more details on the grantmaking process. A wide range of projects are eligible for this RFP. Types of applicants: We’re open to funding both nonprofit and for-profit organizations (including academic institutions), as well as individuals.
However, we prefer nonprofit organizations. This means that if your organization is for-profit or if you apply for individual funding, your project will have to clear a higher bar. Experience: We’re open to funding people at all levels of seniority and organizations at all levels of maturity, but we will value relevant experience highly.
That said, relevant experience can come in many forms, and we especially encourage people without traditional academic credentials to apply. Location: There are in principle no restrictions with respect to location — you can apply from anywhere in the world. However, some locations pose legal and logistical challenges that may make it difficult to fund you.
While we will do our best to fund all projects that excite us, this may be impossible in practice. Available funding: We plan to make grants totaling around $8-10 million for proposals received until February 6, 2026, and expect that in this period, the typical successful proposal will receive funding in the range of $100,000 to $1 million for a project duration between 6 months and 2 years.
What makes for a strong submission When evaluating submissions, we will primarily evaluate (i) how likely the project is to be successful, and (ii) how much impact it could have on high-stakes decision-making if successfully completed. We’ll be more confident about the success of your project if: The budget and project plan are realistic given the project’s goals. You and your team have relevant expertise and/or experience.
For instance: A technical background in machine learning — whether it involves generative AI or more traditional approaches. We expect that some of the most exciting technical projects we fund will combine traditional and gen-AI approaches. A demonstrated track record of developing or marketing apps/products that have found a large user base.
A relevant publication record or academic background (e.g. a Ph. D. in statistics, cognitive science, decision science).
Experience with judgmental forecasting, e.g. experience as a forecaster or experience working for a forecasting platform. If you have relevant expertise or experience that isn’t directly obvious from your CV, we recommend you highlight it elsewhere in your application.
Similar projects have previously been successful (for instance, you could show that the training or post-training approach you’re proposing has previously led to improvements on a task that is structurally similar to forecasting).
The following criteria will make us more confident about the project’s potential impact, though successful submissions won’t need to satisfy all of these (especially #6 and #7): Your submission clearly articulates how your approach will make forecasting more accurate or relevant (AI for probabilistic forecasting), or how it will improve sound reasoning relevant for high-stakes decision-making in a consequential way (AI for sound reasoning).
Your submission includes a description of how your approach differs from the current standard in the industry or research area your project falls into, and why you expect it to have an advantage over alternatives. You/your organization have experience working with high-stakes decision-makers , or you can provide evidence of demand for this project from such decision-makers.
Your submission is focused on reducing global catastrophic risks or provides a clear story for how your idea can be applied to this domain. We understand that you don’t have much space to provide detail in the first submission round, so we only expect you to describe the rough outline of how your project fits these criteria. We will ask for additional detail if your project advances to the next stage.
Further detail on our areas of interest We are looking for proposals for AI models that help to make forecasts more accurate and relevant. We are thinking primarily of forecasting that is Probabilistic: Quantitative forecasts that express uncertainty in terms of probabilities (though this might include the use of confidence intervals over continuous parameters).
Probabilistic forecasts stand in contrast to qualitative forecasting methods like scenario analysis or expert interviews. Judgmental: Forecasts that require judgment-calls by the forecaster, i.e. forecasts in a setting that does not allow straightforward derivation of forecasts via statistical models or trend extrapolation.
This is required in decision-making facing new and untested technology, or when trying to estimate the risk of extreme events that have never occurred, like nuclear war. Judgmental forecasts stand in contrast to model-based quantitative forecasting as it is used in e.g. weather forecasting, supply-chain analysis, or financial trading, which usually require large sets of structured data in the domain of application.
[1] 1 Note that automated forecasters also require such datasets in training, but they do not require them after deployment . Close That said, we see potential in combining probabilistic forecasts with qualitative approaches, and in combining judgmental forecasting with empirically informed quantitative modeling.
Projects in this area that we currently see as most promising mostly fall into two categories: Projects that work directly on automated forecasting , and projects that automate the “infrastructure” that connects forecasting to decision processes . We are also interested in evaluations and benchmarks , although we expect that they will take a smaller share of our funding in this area than in the AI for sound reasoning area.
We are most excited about projects that make AI forecasting (or AI-supported human forecasting) more: Accurate. Here, we’re thinking of accuracy in terms of how forecasts are assessed by a proper scoring rule (see Gneiting and Raftery , 2007, for an academic overview). Proper scores combine an evaluation of a forecast’s informativeness (did it assign high probability to events that actually occurred?)
, and calibration (did events assigned a probability 70% occur roughly 70% of the time?) , and can sometimes be decomposed precisely into these two components. Relevant.
The forecasting questions are relevant for high-stakes decision-making , e.g., they are consequential for a large number of individuals. They should also be action-guiding . For instance, the probability of a candidate winning an election moving from 45% to 60%, or the predicted average global temperature in 10 years increasing by 0.
1°C, won’t have clear policy implications for most individual actors, even if the forecast events themselves are consequential. Forecasts of conditional outcomes (if a carbon tax of X% is implemented, emission levels are estimated to be Y), and collections of forecasts of more fine-grained events (e.g. hurricane incidence in specific regions) will often be more relevant in that sense.
Examples of projects we’re interested in funding AI forecasters: Systems that can respond to forecasting questions as they might appear on a public forecasting platform. Typically, forecasting platforms ask for probabilities for discrete outcomes or for point estimates of a continuous variable.
We’re interested both in AI forecasters that deal with these kinds of questions and in those that handle more complex questions — for instance, systems that output a full forecasting model involving explicit causal relationships between quantitative variables. Existing work in this area includes Schoenegger et al (2023) , Halawi et al (2024) , FutureSearch , and submissions to Metaculus’ AI Benchmark competition .
The key metric for such systems will typically be accuracy . AI forecasters are currently nearing performance parity with pooled non-expert human forecasts, and we anticipate they will soon match or outperform top superforecasters (see our Q&A for why we believe this). This would make AI forecasting significantly cheaper, faster, and easier to deploy than human forecasting, enabling broader application in critical decision-making.
Forecasting model builders: Quantitative forecasting models (e.g. causal influence diagrams or programs in a probabilistic programming language) are sometimes an important analytical tool that improves the performance of human forecasters. They also make forecasters’ reasoning more transparent to other forecasters, and to stakeholders who ultimately make use of the forecasts and rationales they produce.
Developing AI tools that assist or automate the development of such models (perhaps after seeing a prompt with a qualitative explanation of a given model) could increase the performance and transparency of human and AI forecasters. We’re particularly interested in tools designed with AI forecasters in mind, given that we expect a growing share of forecasting work to be done by AIs in the future.
Existing work in this domain includes SquiggleAI , an LLM-powered app that generates probabilistic programs for estimation. Model-building might improve forecasting on several dimensions, but will especially make it more relevant .
Automated foresight or risk analysis: Another important class of tools could generate qualitative foresight analyses complementary to quantitative forecasting, in particular scenario planning : the concrete description of possible future events for the purpose of informing decision-making. For instance, users might start by describing the features of the future they’re envisioning and some criteria the scenario should satisfy.
Then, the tool will suggest specific events or decision points that the scenario could contain. The tool might also run a simulation of the scenario (often called a “ wargame ”), though we’re unsure whether current AI capabilities will be sufficient for coherent real-time gameplay.
Automated question generation for training and evaluating AI forecasters: One approach to building better AI forecasters involves generating and resolving new forecasting questions. Currently, public forecasting platforms offer a fairly limited set of questions, many of which have already been resolved (which makes it harder to use them to evaluate forecasting skill ).
Thus, generating a large and curated set of forecasting questions that resolve in the near-term future will likely improve the training process for AI forecasters. We’d be interested to see research into validating the usefulness of these question sets for increasing the accuracy of AI forecasters.
Connecting Forecasting to Decision Processes Using forecasting for decision-making requires “infrastructure” that connects the forecasts to the processes that determine the decision. A key part of such infrastructure will be to ensure that forecasts are addressing the right questions, and that forecasts are legible to decision-makers.
Automated forecasting rationales: Forecasting rationales are explanations of the reasoning underlying a forecast, similar to a forecasting model, but typically of a more qualitative nature. Improving rationale quality might be an effective way to build better AI forecasters.
Rationales can also make AI forecasters more legible, as a pure accuracy track record is often not sufficient to convince decision-makers (see also related commentary by FutureSearch). Rationales can also allow easier combination of forecasts, including with forecasts that use other sources of information, and they can be tailored to the needs of decision-makers with different backgrounds and idiosyncrasies.
Automated question decomposition: Question decomposition is central to defining forecastable proxies for high-level, decision-relevant questions that are otherwise too vague to operationalize. A question decomposition tool would support this by taking ambiguous prompts like “What will be the overall effect of X? ” and breaking them down into subquestions about measurable, forecastable quantities.
This often involves identifying intermediate outcomes, precursors, or necessary conditions that can be forecast directly. In the context of global catastrophic risks, where direct feedback is rare or delayed, this decomposition is especially critical. [2] While there has already been some research on this, we think AI has significant potential for expanding and accelerating it.
Question decomposition will likely help with making forecasting more relevant . 2 For example, trying to forecast the odds of an engineered pandemic is more complicated in some ways than forecasting the odds of an economic recession; we have a lot of historical data about recessions, and practically none about engineered pandemics.
So rather than looking at past data, we need to decompose questions about engineered pandemics into sub-questions that give us a way to use what context we have; for example, we could create subquestions about progress in DNA synthesis technology or the number of active terrorist groups with access to certain resources.
Close Evaluations and benchmarks Evaluations and benchmarks: We’re further interested in benchmarks that evaluate how relevant AI forecasting tools are, including tools for question generation.
We’re also open to funding further evaluations for forecasting accuracy , but submissions would need to present a strong case for why they would be improvements over existing efforts like ForecastBench , Bench to the Future , and the Metaculus AI Benchmark .
AI models, especially LLMs (and LLM-based models) are being incorporated into decision-making processes at a rapid pace across society, including in legal, academic and political domains. We expect this trend to continue over the coming years, and potentially to accelerate.
It seems crucial to us that models that are used for highly consequential decisions follow principles of sound reasoning, and nudge your users towards following those principles as well. By default, AI companies are strongly incentivized to create broadly capable, agentic models that can accomplish real-world tasks.
In significant part, we expect them to do this by iteratively trying to improve performance on easy-to-measure tasks of clear economic importance. This will improve their ability to solve a wide variety of epistemic challenges, turning models into general-purpose digital assistants.
They’ll likely integrate search and reasoning features, learn to perform tasks in unfamiliar domains without needing extensive retraining or human instruction (“learning on the job”), and excel at tasks like navigating browsers or APIs. We want to support work that takes increasing capabilities in such domains as a backdrop, and adds on top something that AI companies may not provide by default.
Typically, we expect that AI companies will have an incentive for their models not to commit egregious factual errors , and will typically have an incentive to implement high standards with respect to a model’s core function.
Thus, we expect AI companies to reduce the rate of hallucinations / confabulations by default, and also expect that specific tools like research assistants will become more reliable, and will generally not make up academic papers or court cases when the technology has matured.
However, we believe that there will still be many more subtle errors and biases that won’t hurt AI models’ commercial viability, but might have strongly adverse effects on the reliability of their reasoning.
We want to especially flag two directions: Research: We are interested in evaluations of AIs with respect to principles of sound reasoning like truthfulness, metareasoning, or consistency (see below), and also in funding work that investigates the causes of AI displaying or not displaying these principles, and in work that develops techniques or builds relevant datasets for AI to display them more reliably.
Developing AI tools that directly help with specific tasks that are disproportionately helpful for sound reasoning, like fact-checkers, arbitrators, argument analyzers, or fact tracers.
We believe that evaluations and related research matter not only because they help researchers improve on the relevant metrics, but also because they might create a race to the top among AI developers to improve their models on the evaluated characteristics. Therefore, we’d like the evaluation results to be easily accessible (e.g. on a dedicated website) and updated to include new models as they are developed.
We will preferentially (but not exclusively) fund efforts that plan to do this, and are also potentially interested in projects entirely focused on curation and communication of existing benchmarks.
Examples of projects we’re interested in funding Here we list examples of principles that seem relevant for sound reasoning and that could plausibly satisfy the above criteria: Truthfulness: Seeking truthful sources and making truthful statements is a core epistemic virtue, and there are several promising ways to operationalize it: Detection of false or misleading claims.
The ability to identify incorrect or deceptive statements in longer contexts, including against adversarial pressure (i.e. in the context of texts deliberately meant to mislead). Informativeness. How much better informed (if at all) are humans after interacting with an AI model?
We see room for several experimental studies in this area, across different domains of knowledge and types of interactions. There are also several promising tools in this area like fact-checkers and fact-tracers, i.e. tools that start with a claim, search for sources, and recursively follow them.
They might then output a judgment on whether the original claim can be supported (fact-checkers), or produce a report that provides context on the root-source of a claim (fact-tracers), e.g. tracing a common belief to its first mention, identifying the original/seminal work in a line of research, identifying the primary source for a quote, etc. Tools applying the concept of Community Notes to context-provision in other domains also seem promising.
Meta-reasoning: we expect that the AI models most useful for sound reasoning will be highly aware of their own limitations and will be able to reason competently about their own reasoning. [3] Research in this area might measure one of the following: Calibration: Does the confidence AI models express in their own claims match actual accuracy?
One potential avenue is to measure confidence in probabilistic terms, evaluated analogously to probabilistic forecasts. However, confidence could also be expressed in different ways, e.g. through the willingness or refusal to answer a prompt, through deference to other sources, through asking the user clarifying questions, etc., and we’d also be interested in proposals for how to quantify and evaluate those.
Explanation: Evaluating whether models can accurately explain their own output. We might for instance test this by asking a model to justify an answer, then testing whether changing the factor it cites leads to a different answer. This method helps to assess whether stated rationales are genuine drivers of decisions.
This method is related to the study of chain-of-thought faithfulness, but is distinct from this area in that the rationales would be generated after the original output is produced. Self-awareness: Are models aware of their own biases and inconsistencies? Can they predict whether they’ll be biased or inconsistent on a given question or topic?
Systems with strong self-awareness could provide users with more reliable guidance about when to trust their responses. Consistency: An important characteristic of epistemically sound AI models is that their output should be internally consistent, robust across users, and not vulnerable to framing effects (giving answers that match “desired” responses implicit in the prompt).
Promising properties in this cluster are: Vulnerability to framing: How much does the output generated by the model depend on framing effects? Is it possible to move responses in a particular direction (e.g. a particular political ideology) through subtle variations in otherwise semantically equivalent prompts?
Sycophancy: Do language models systematically agree with or flatter users, especially when presented with controversial or value-laden topics?
There is a reasonably large existing literature on this , but sycophancy seems to be a case where market forces will consistently have adversarial incentives (since sycophancy, at least when mild, will often lead to larger user satisfaction) and can be particularly damaging for sound reasoning, so more work on careful measurements in this area still seems very valuable.
Logical consistency: Does the model’s output adhere to (logical) consistency requirements? [4] Tests could measure understanding logical relationships between different concepts (e.g. understanding that properties like height are transitive), adhering to probability laws (e.g. not falling prey to the conjunction fallacy ), or the frequency of direct self-contradiction.
Navigating debates: Sound reasoning often involves engaging with contradictory evidence and opposing viewpoints. Research in this area could evaluate how models update (or don’t) in response to counterarguments or new evidence, or whether they can faithfully reproduce different viewpoints on a controversial question.
This principle of sound reasoning also is related to argument analyzers , a broad class of potential AI tools that analyze the structure of human arguments and suggest improvements, addressing several principles at once. Such tools could break down a text written by a human into different components (e.g. premises, conclusions, and the relationship between them), and spot errors, biases, and inconsistencies.
Other versions of such tools could suggest counterarguments, or ways of strengthening the existing argument. Another class of tool related to this principle would be arbitrators , tools that help different sides of a discourse to communicate with each other, one variant of which would be a disagree-and-bet operationalizer .
Closely related to automated question generation in the forecasting context, a disagree-and-bet operationalizer reads through a discussion between people with opposing views (e.g. a debate transcript, or a series of articles replying to each other), and suggests bets the participants could make based on their views.
3 Assigning mental states like “awareness” or competencies like “reasoning” to current AI models is controversial, but we mean it here in the restricted sense of AI models highly reliably behaving as if they possessed these states or competencies, in the relevant domain of application. Close 4 See here for early research on this.
Close Please note that the properties and tools listed above represent our current best guess about the most exciting work in this area. However, we are much more uncertain about this area than we are about AI for forecasting, and would not be surprised if some of the best projects we end up funding were not to be found on this example list.
To evaluate whether a project idea of yours would be a good fit for this RFP, you should evaluate whether it fits our background motivation, and consult our submission guide . What kinds of progress are you expecting from AI forecasters? h2+p]:relative"> We think AI forecasting models offer a scalable and rapidly improving opportunity.
They are currently nearing performance parity with pooled non-expert human forecasts, [5] and we anticipate they will soon match or outperform top superforecasters. This would make AI forecasting significantly cheaper, faster, and easier to deploy than human forecasting, enabling broader application in critical decision-making.
5 Some evidence for this includes the ForecastBench leaderboard (where AI models do slightly worse than superforecasters but outperform non-expert public participants in terms of Brier score as of Oct 8, 2025 ) and the Metaculus AI Benchmarking Tournament (where Metaculus Pro forecasters outperformed AI models in terms of log score in all three quarters between 2024 Q3 and 2025 Q1, although the 2024 Q4 difference didn’t reach the conventional threshold for statistical significance).
Close Scalability: AI models offer cheap, on-demand forecasting, free from some human limitations like fatigue, bias, or the risk of leaking sensitive information. [6] They can be integrated into tools and adapted to institutional needs, and can handle multiple questions at once.
Even without full automation, models can support individual forecasting steps — like decomposing questions, finding evidence, fitting models, or writing rationales. Once they reach human level, their scalability and availability will make them even more useful. [7] Fast improvement: We also believe that the performance of AI forecasters, in particular LLM-based forecasters, is likely to improve significantly.
Several studies and public benchmarks already suggest that GPT-4o-level models match the performance of average human forecasters across a range of domains (while still falling short of top-human performance), [8] and there’s evidence of a consistent scaling trend: forecasting accuracy improves with training compute (see Figure 1).
If this continues, frontier models might reach superforecaster-level performance before 2030 — even without bespoke fine-tuning or specialized forecasting architectures. [9] 6 While we acknowledge that AI models introduce biases and privacy risks of their own, we believe that these are easier to monitor and mitigate than in humans. Close 7 This argument rests on a few core assumptions: that inference costs of AI models will continue
Scoring criteria used to review proposals for this grant.
According to the current listing, eligibility includes: Academic researchers, independent scholars, and nonprofits. Confirm the full requirements in the official notice before applying.
The current listing shows $100,000 - $1,000,000. Verify award ceilings, matching requirements, and allowable costs in the official notice.
AI for Forecasting and Sound Reasoning is funded by Coefficient Giving (formerly Open Philanthropy). Verify program details on the funder's official page before applying.
This listing is flagged as international in scope. Check the official notice for country-specific restrictions before applying.
Start from the official opportunity page linked in this listing — it carries the sponsor's submission instructions.
Coefficient Giving (formerly Open Philanthropy) is a grant from Coefficient Giving that funds evidence-based programs and research in global health, animal welfare, scientific research, and effective altruism causes where philanthropic dollars can achieve high impact per dollar. The fund supports organizations demonstrating rigorous evidence of effectiveness and scalable potential. Eligible applicants include nonprofits, research institutions, and projects aligned with Coefficient Giving's priority cause areas. The fund emphasizes transparency, cost-effectiveness analysis, and funding gaps not addressed by government or traditional philanthropy.
Request for proposals: Improving Global Health and Wellbeing in an Era of Transformative AI is sponsored by Coefficient Giving (formerly Open Philanthropy) - Global Health and Wellbeing Opportunities Fund. Supports research, policy, implementation, and field-building to help ensure AI improves human health and wellbeing, with particular attention to risks for and benefits to the global poor.
Brown Girl Jane x SheaMoisture Grant is a grant from SheaMoisture and Brown Girl Jane that funds Black and woman-owned beauty and wellness businesses in the United States. Part of SheaMoisture's broader commitment to addressing racial inequality through its $1 million annual giving fund, this program specifically supports founders at the intersection of Black and women-owned entrepreneurship in the beauty and wellness sector. Applicants must be based in the U.S. and have operated their business for at least one year. Grants range from $10,000 to $25,000. Check the SheaMoisture Fund website for the current open cycle, as deadlines vary by cohort.
Support Adoption Grant Program is sponsored by Texas Office of the Attorney General. This program provides critical resources for pregnant women considering adoption and support for children awaiting placement with adoptive parents. Purpose areas include material needs for pregnant women, needs of children awaiting placement, training and advertising related to adoption, and pre- and post-adoption counseling.