ai4bharat/romansetu-cpt-native-200m

TEXT GENERATIONPricing:Input $0.4 / Cached $0.08 / Output $0.8Concurrent Unit Cost:1Model Size:7BQuant:FP8Context Size:4kPublished:Mar 7, 2025License:llama2Architecture:Transformer Open Weights Featherless Exclusive Cold

ai4bharat/romansetu-cpt-native-200m is a 200 million parameter causal language model developed by AI4Bharat. This model is specifically designed for efficient multilingual capabilities through Romanization, as detailed in the RomanSetu research paper. It focuses on unlocking language model potential across various languages by leveraging Roman script. Its primary use case is enabling multilingual text processing and generation via Romanized input.

Loading preview...

Model Overview

The ai4bharat/romansetu-cpt-native-200m is a 200 million parameter causal language model developed by AI4Bharat. It is a key component of the research presented in the paper "RomanSetu: Efficiently unlocking multilingual capabilities of Large Language Models via Romanization" (arxiv.org/abs/2401.14280). This model is engineered to enhance the multilingual capabilities of large language models by processing Romanized text.

Key Capabilities

  • Multilingual Processing: Designed to handle various languages through their Romanized forms.
  • Efficient Language Unlocking: Aims to make LLMs more accessible and effective for diverse linguistic inputs.
  • Research-Backed: Developed as part of a specific research initiative focusing on Romanization techniques.

Good For

  • Multilingual Applications: Ideal for scenarios requiring text processing or generation across multiple languages using Romanized input.
  • Research and Development: Suitable for researchers exploring Romanization strategies in natural language processing.
  • Low-Resource Languages: Potentially beneficial for languages with limited native script data by leveraging Romanized representations.