RKB109/audio-event-triage-20260724-dataset
Viewer • Updated • 18 • 44
This repository contains a small, transparent prototype model for Operations teams need an explainable starting point for classifying alarms, machinery noise, and speech-like events.
The model combines per-label token weights with IDF-weighted evidence retrieval. It was generated for reproducible architecture demonstrations and does not call a hosted LLM.
audio-classificationautomatic-speech-recognitionfeature-extractionaudio-to-audioThe included records are synthetic feature vectors and do not replace evaluation on licensed real audio.
The dataset is synthetic and small. Do not use this model for consequential decisions without representative data, expert review, and production-grade evaluation.
The linked GitHub repository includes train.py, the exact dataset split,
evaluation code, and the model JSON format.