Instructions to use nikhilchandak/OpenForecaster-8B with libraries, inference providers, notebooks, and local apps. Follow these links to get started.

Libraries

How to use nikhilchandak/OpenForecaster-8B with Transformers:

# Use a pipeline as a high-level helper
from transformers import pipeline

pipe = pipeline("text-generation", model="nikhilchandak/OpenForecaster-8B")
messages = [
    {"role": "user", "content": "Who are you?"},
]
pipe(messages)

# Load model directly
from transformers import AutoTokenizer, AutoModelForCausalLM

tokenizer = AutoTokenizer.from_pretrained("nikhilchandak/OpenForecaster-8B")
model = AutoModelForCausalLM.from_pretrained("nikhilchandak/OpenForecaster-8B")
messages = [
    {"role": "user", "content": "Who are you?"},
]
inputs = tokenizer.apply_chat_template(
	messages,
	add_generation_prompt=True,
	tokenize=True,
	return_dict=True,
	return_tensors="pt",
).to(model.device)

outputs = model.generate(**inputs, max_new_tokens=40)
print(tokenizer.decode(outputs[0][inputs["input_ids"].shape[-1]:]))

Notebooks
Google Colab
Kaggle
Local Apps Settings

vLLM

How to use nikhilchandak/OpenForecaster-8B with vLLM:

Install from pip and serve model

# Install vLLM from pip:
pip install vllm
# Start the vLLM server:
vllm serve "nikhilchandak/OpenForecaster-8B"
# Call the server using curl (OpenAI-compatible API):
curl -X POST "http://localhost:8000/v1/chat/completions" \
	-H "Content-Type: application/json" \
	--data '{
		"model": "nikhilchandak/OpenForecaster-8B",
		"messages": [
			{
				"role": "user",
				"content": "What is the capital of France?"
			}
		]
	}'

Use Docker

docker model run hf.co/nikhilchandak/OpenForecaster-8B

SGLang

How to use nikhilchandak/OpenForecaster-8B with SGLang:

Install from pip and serve model

# Install SGLang from pip:
pip install sglang
# Start the SGLang server:
python3 -m sglang.launch_server \
    --model-path "nikhilchandak/OpenForecaster-8B" \
    --host 0.0.0.0 \
    --port 30000
# Call the server using curl (OpenAI-compatible API):
curl -X POST "http://localhost:30000/v1/chat/completions" \
	-H "Content-Type: application/json" \
	--data '{
		"model": "nikhilchandak/OpenForecaster-8B",
		"messages": [
			{
				"role": "user",
				"content": "What is the capital of France?"
			}
		]
	}'

Use Docker images

docker run --gpus all \
    --shm-size 32g \
    -p 30000:30000 \
    -v ~/.cache/huggingface:/root/.cache/huggingface \
    --env "HF_TOKEN=<secret>" \
    --ipc=host \
    lmsysorg/sglang:latest \
    python3 -m sglang.launch_server \
        --model-path "nikhilchandak/OpenForecaster-8B" \
        --host 0.0.0.0 \
        --port 30000
# Call the server using curl (OpenAI-compatible API):
curl -X POST "http://localhost:30000/v1/chat/completions" \
	-H "Content-Type: application/json" \
	--data '{
		"model": "nikhilchandak/OpenForecaster-8B",
		"messages": [
			{
				"role": "user",
				"content": "What is the capital of France?"
			}
		]
	}'

Docker Model Runner
How to use nikhilchandak/OpenForecaster-8B with Docker Model Runner:
```
docker model run hf.co/nikhilchandak/OpenForecaster-8B
```
Browse Quantizations to use this model in llama.cpp, Ollama, LM Studio, or any compatible app.

OpenForecaster-8B

OpenForecaster-8B is a specialized language model for open-ended forecasting and predicting future events. This model is post-trained from Qwen3-8B using reinforcement learning on the OpenForesight dataset. It was introduced in the paper Scaling Open-Ended Reasoning to Predict the Future.

Performance of OpenForecaster-8B on FutureX

Performance on FutureX benchmark in July-August 2025 on non-numeric questions (86 Qs): OpenForecaster-8B has a much higher accuracy than 100B+ models. We limit to models released before April 2025 for a fair, equal knowledge cutoff comparison.

Model Description

OpenForecaster-8B is trained to make calibrated predictions on open-ended questions about future events. The model has been trained to:

Provide calibrated confidence estimates when asked (please prompt explicitly)
Reason about uncertainty and future scenarios
Leverage retrieved information (when provided in context) to improve predictions

Note: OpenForecaster-8B's knowledge cutoff is, at best, till April 2025 (base model's cutoff being ~June 2024) so it has no knowledge about the events that have happened since then till now. Thus, if you ask it questions about 2026 or later without providing recent developments/relevant context, it will only be able to answer from its parametric knowledge which might not be helpful/up-to date. Thus, please be aware of this and use it with RAG over recent developments if possible.

Training

This model was trained on the OpenForesight dataset, which contains over 52,000 forecasting questions generated from global news events. The training was done using GRPO optimizing a joint reward function combining accuracy and brier score. Please check the paper for more details.

Base Model: Qwen3-8B
Training Dataset: OpenForesight

Usage

from transformers import AutoTokenizer, AutoModelForCausalLM

model_name = "nikhilchandak/OpenForecaster-8B"
tokenizer = AutoTokenizer.from_pretrained(model_name)
model = AutoModelForCausalLM.from_pretrained(model_name)

# template 
prompt = "What is the likelihood that [future event] will occur by [date]?"
# example
prompt = "Who will become the next Prime Minister of India based on the general election to be held in 2029? Provide specific predictions with probabilities."
inputs = tokenizer(prompt, return_tensors="pt")
outputs = model.generate(**inputs, max_length=8192)
prediction = tokenizer.decode(outputs[0], skip_special_tokens=True)
print(prediction)

Performance

OpenForecaster-8B achieves competitive performance with much larger models like DeepSeek-v3 and Qwen3-235B-A22B on forecasting benchmarks. Key improvements include:

Improved Accuracy: Better prediction of future events
Better Calibration: More reliable confidence estimates
Enhanced Consistency: Reduced logical violations in predictions

Citation

If you use this model in any way, please cite the corresponding paper:

@article{chandak2025scaling,
 title={Scaling Open-Ended Reasoning to Predict the Future},
 author={Chandak, Nikhil and Goel, Shashwat and Prabhu, Ameya and Hardt, Moritz and Geiping, Jonas},
 journal={arXiv preprint arXiv:2512.25070},
 year={2025}
}