My Bilingual LLM

Summary

My Bilingual LLM is a compact decoder-only Transformer language model designed for experimental Bangla and English text generation.

Architecture

  • Parameters: approximately 41.87M
  • Vocabulary size: 32,000
  • Embedding dimension: 512
  • Transformer layers: 8
  • Attention heads: 8
  • Maximum context length: 512
  • Weight tying: enabled

Languages

  • Bangla
  • English
  • Mixed Bangla-English input

Intended Use

This model is intended primarily for:

  • research
  • experimentation
  • educational purposes
  • portfolio demonstration
  • language-model inference experiments

Limitations

This is an experimental model and is not production-ready.

Known limitations include:

  • repetitive generation
  • weak long-form coherence
  • occasional mixed-language output
  • incorrect factual statements
  • unusual numeric sequences
  • grammatical errors

Generation

  • temperature: 0.78
  • top_k: 50
  • top_p: 0.92
  • repetition_penalty: 1.12
  • no_repeat_ngram_size: 3
  • max_new_tokens: 80

Checkpoint Integrity

Model SHA256:

0acef793ab3449d7a165c6d2a840297ad45a477344fe9064147b1f19e5ae6c1d

Tokenizer SHA256:

f9e19f1b7169fcd243a433a3448181bffb6a6e23d0cdf34909100d5225e74cc5

License

Verify the licenses of the training data, tokenizer, source code, and dependencies before redistribution.

Downloads last month
-
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support