Team Ai
Modelpublic

Itbanque/whisper-ja-zh-tiny

sourceHugging Facecc-by-4.0updated 1y agoView on Hugging Face
0likes67downloads
Model Card

Whisper JA-ZH Tiny

A fine-tuned OpenAI Whisper tiny model on Japanese-to-Chinese speech translation, trained on a subset of the DataLabX/ScreenTalk_JA2ZH dataset.


๐Ÿ“Œ Model Details

  • โ€”Base model: openai/whisper-tiny
  • โ€”Task: Speech translation (Japanese โ†’ Chinese)
  • โ€”Dataset: ScreenTalk-JA2ZH (private subset)
  • โ€”Training framework: ๐Ÿค— Transformers + Seq2SeqTrainer
  • โ€”Hardware: RTX 5090
  • โ€”Mixed Precision: FP16 enabled
  • โ€”Total Training Epochs: Early-stopped at 11 epochs
  • โ€”Eval BLEU: 0.757 on held-out eval set, 0.609 on held-out test set.

๐Ÿƒ Training Configuration

yaml
train_batch_size: 96
eval_batch_size: 64
learning_rate: 3e-4
warmup_steps: 1000
num_train_epochs: 20
gradient_accumulation_steps: 1
save_steps: 1000
eval_steps: 1000
logging_steps: 1000
fp16: true
eval_strategy: step
early_stopping: enabled (patience=5)
Best checkpoint auto-loaded via load_best_model_at_end=True using eval_bleu as the metric.

๐Ÿ“ˆ Test Dataset

Final run metrics (test set):

loss: 2.3245
bleu: 0.6095

๐Ÿ“ Structure

Repository includes:

  • โ€”config.json, generation_config.json, preprocessor_config.json
  • โ€”Tokenizer: tokenizer_config.json, vocab.json, merges.txt, etc.
  • โ€”Training log: training_20250610-194336.log
  • โ€”TensorBoard logs: runs/

๐Ÿš€ How to Use

python
from transformers import WhisperProcessor, WhisperForConditionalGeneration

processor = WhisperProcessor.from_pretrained("fj11/whisper-ja-zh-tiny")
model = WhisperForConditionalGeneration.from_pretrained("fj11/whisper-ja-zh-tiny")

๐Ÿ“ฌ Contact

For business inquiries or collaboration, visit https://www.itbanque.com or reach out via Hugging Face.


๐Ÿ“œ License

CC BY-NC-SA 4.0 (Non-commercial, Attribution, ShareAlike)