m3rg-iitd
/

llamat-3

materials science

large language model

Model card Files Files and versions Community

Model Card for LLaMat-3

LLaMat-3 is a specialized large language model designed to be a foundational large language model for materials science.

Overview

Model Type: Large Language Model (LLM)
Base Model: LLaMat-3 (continued pretraining of LLaMA-3 on material science data)
Language: English
License: LLaMA-3 License
Tags: Material Science, Domain Adaptation, Table Understanding, Scientific Data Parsing, Materials Copilot

Model Details

Key Features

Applications: Can be finetuned for information extraction, table understanding, parsing data for research tasks, and crystal structure generation.

Development and Support

Developed by: M3RG, IIT Delhi & DAIR, IIT Delhi
Compute Support:
- Edinburgh International Data Facility (EIDF): Provided access to Cerebras CS2 clusters for pretraining.
- IIT Delhi High-Performance Computing Cluster: Supported fine-tuning and inference stages.

Technical Specifications

Hardware Infrastructure

Pretraining: 2 Cerebras CS-2 Wafer-Scale Engines (WSE-2)

Software Stack

Frameworks: PyTorch, Hugging Face Transformers

Model Sources

Repository: LLaMat on GitHub
Compute Resources: EIDF Cerebras CS Clusters

Downloads last month: 6

Safetensors

Model size

8.03B params

Tensor type

F32

·

Inference API

Unable to determine this model's library. Check the docs .

Model tree for m3rg-iitd/llamat-3

Base model

meta-llama/Meta-Llama-3-8B

Finetuned

(374)

this model

Finetunes

Collection including m3rg-iitd/llamat-3

LLaMat

Foundational Large Language Models for Materials Research • 6 items • Updated 12 days ago • 3