Skip to main content

Minerva: Unleashing the Secrets of advanced Mathematics 🏛️🔢

Minerva is a groundbreaking language model that pushes the boundaries of mathematical understanding and problem-solving. Designed with an advanced math theme, Minerva embodies the spirit of renowned mathematicians such as Euclid, Pythagoras, and Archimedes. By harnessing their advanced wisdom, Minerva offers unparalleled capabilities in mathematical reasoning and exploration.


GitHub issues GitHub forks GitHub stars GitHub license

Share on Twitter Share on Facebook Share on LinkedIn

Share on Reddit Share on Hacker News Share on Pinterest Share on WhatsApp


Install

pip install minerva

Usage

import torch
from minerva import Minerva, Train

# Example usage
x = torch.randint(0, 20000, (1, 1024))

Minerva(x)

# or train
Train()

Training

To train Minerva, follow these steps:

  1. Configure the training settings by setting the environment variables:

    • ENTITY_NAME: Your wandb project name
    • OUTPUT_DIR: Specify the output directory for saving the weights (e.g., ./weights)
  2. Launch the training process using Deepspeed:

Accelerate Config
Accelerate launch train_distributed_accelerate.py

Dataset Building

To build a custom dataset for Minerva, you can preprocess the data using the build_dataset.py script. This script performs tasks such as pre-tokenization, data chunking, and uploading to the Huggingface hub. Here's an example command:

Dataset Description
Mathematical Web Pages Web pages containing mathematical expressions in MathJax format, cleaned to preserve math notation
arXiv 2 million arXiv papers up to Feb 2021, in LaTeX format
General Natural Language Data Same dataset used to pretrain PaLM models

The mathematical web pages and arXiv datasets focus on technical and mathematical content. The general natural language data provides a broad coverage of general language.

The paper states the mathematical web pages and arXiv each account for 47.5% of the total data. The remaining 5% is general natural language data which is a subset of what was used for PaLM pretraining.

Roadmap 🗺️📍

  • Create a dataset of ARXVIV papers

Metadata

Release files for minerva-torch 0.0.1

For a detailed explanation of source distributions (sdists) and built distributions (wheels), please see the package formats documentation.

Source distribution (sdist)

Source distribution for minerva-torch 0.0.1
File Size Uploaded
minerva_torch-0.0.1.tar.gz 15.4 kB Details

Built distribution (wheel)

Table of built distributions (wheels) for minerva-torch 0.0.1
File Interpreter ABI Platform
minerva_torch-0.0.1-py3-none-any.whl Python 3 none any Details

Total release size: 30.5 kB

Release files / minerva_torch-0.0.1.tar.gz

Download URL minerva_torch-0.0.1.tar.gz
Size 15.4 kB
Tags Source
SHA-256 checksum
How to use checksums
27941333b7b0bcc1292cdf86864c4d192d52d1f498d6716c45206f8388ce6155
BLAKE2b-256 checksum
How to use checksums
0285b2c35adb9f9147c40fa1dda5c9eb4cf511d441a1f659564ae3b65018a2a3
Upload date
Uploaded using Trusted Publishing?
What is trusted publishing?
No
Uploaded via poetry/1.3.2 CPython/3.11.0 Darwin/22.4.0

Release files / minerva_torch-0.0.1-py3-none-any.whl

Download URL minerva_torch-0.0.1-py3-none-any.whl
Size 15.1 kB
Tags Python 3
SHA-256 checksum
How to use checksums
5f9e67cc4c424c93ef06bf4da12b76d0fa210f308b31e6c2391096a7baa762d0
BLAKE2b-256 checksum
How to use checksums
f33d3470d7082580683ebf6c12798d910105f0d4db38558403b4d3f95ac35edf
Upload date
Uploaded using Trusted Publishing?
What is trusted publishing?
No
Uploaded via poetry/1.3.2 CPython/3.11.0 Darwin/22.4.0

Release history Release notifications | RSS feed

This release

0.0.1 This release

2 release files

Anthropic, PBC Visionary sponsor Bloomberg Visionary sponsor Hudson River Trading Visionary sponsor Meta Visionary sponsor NVIDIA Visionary sponsor Microsoft Sustainability sponsor Depot Continuous Integration AWS Cloud computing and Security Sponsor Datadog Monitoring Fastly CDN Google Download Analytics Sentry Error logging StatusPage Status page