Downloads · 30 days
15
7% of all-time downloads
monodox/arka-1-1b
arka-1-1b is a text generation model from monodox. Use it when you need the model to write or continue text. The card lists the license as apache-2.0.
artistics/arka-1-1b is a multilingual text-generation model for English and Malayalam content in the domain of classical arts and cultural heritage. The model is fine-tuned on a Kerala classical arts corpus, with supp…
Downloads · 30 days
15
7% of all-time downloads
All-time downloads
223
Public
Repo size
112 B
Likes
0
Public
Click a slice to open those files.
.md4 KB · 35%
From the Hugging Face model README
artistics/arka-1-1b is a multilingual text-generation model for English and Malayalam content in the domain of classical arts and cultural heritage. The model is fine-tuned on a Kerala classical arts corpus, with supporting language resources and evaluation files included in this repository.
artistics/arka-1-1bThis repository includes the complete Hugging Face model repository structure. Placeholder scaffold files are provided for model.safetensors, tokenizer.model, and fine-tuning/adapter_model.safetensors; replace them with the trained model, tokenizer, and adapter artifacts before publishing or using the model for inference.
This model is intended for applications involving Kerala classical arts, cultural heritage documentation, educational content generation, multilingual question answering support, and domain-specific writing assistance in English and Malayalam.
Example use cases include:
The model is specialized for classical arts and cultural heritage content, especially related to Kerala. It may be less reliable outside this domain and should not be treated as an authoritative source without human review. Generated content may contain factual errors, omissions, or culturally sensitive inaccuracies.
For educational, archival, or public-facing use, outputs should be reviewed by domain experts.
The model was fine-tuned on a Kerala classical arts corpus. The repository includes language-specific supporting resources:
languages/english/corpus.txtlanguages/malayalam/corpus.txtlanguages/english/vocab.jsonlanguages/malayalam/vocab.jsonFine-tuning artifacts are provided under fine-tuning/:
fine-tuning/adapter_config.jsonfine-tuning/adapter_model.safetensorsEvaluation files are included under eval/:
eval/english-qa.jsonleval/malayalam-qa.jsonlThese files can be used to assess model behavior on English and Malayalam domain-specific question answering examples.
from transformers import AutoModelForCausalLM, AutoTokenizer
model_id = "artistics/arka-1-1b"
tokenizer = AutoTokenizer.from_pretrained(model_id)
model = AutoModelForCausalLM.from_pretrained(model_id)
prompt = "Explain the cultural significance of Kathakali in Kerala."
inputs = tokenizer(prompt, return_tensors="pt")
outputs = model.generate(
**inputs,
max_new_tokens=200,
temperature=0.7,
top_p=0.9,
)
print(tokenizer.decode(outputs[0], skip_special_tokens=True))
.
|-- README.md
|-- LICENSE
|-- .gitattributes
|-- config.json
|-- generation_config.json
|-- tokenizer.json
|-- tokenizer.model
|-- tokenizer_config.json
|-- special_tokens_map.json
|-- model.safetensors
|-- languages/
| |-- english/
| | |-- vocab.json
| | `-- corpus.txt
| `-- malayalam/
| |-- vocab.json
| `-- corpus.txt
|-- fine-tuning/
| |-- adapter_config.json
| `-- adapter_model.safetensors
`-- eval/
|-- english-qa.jsonl
`-- malayalam-qa.jsonl
This model is released under the Apache License 2.0. See the LICENSE file for details.
classical-arts kerala malayalam english multilingual cultural-heritage fine-tuned