ASR/STT subtitle generator. Uses Qwen3-ASR, local LLM, Whisper, TEN-VAD. Noise-robust for JAV
-
Updated
Sep 17, 2026 - Python
ASR/STT subtitle generator. Uses Qwen3-ASR, local LLM, Whisper, TEN-VAD. Noise-robust for JAV
WFGY is heading toward WFGY 5.0 Polaris Protocol, a major open-source release for AI reasoning, RAG, agents, and real-world workflows. Includes Problem Map, Global Debug Card, WFGY 4.0, and the CFV Easter Egg.
Stop your AI from making things up — it proposes, deterministic tools decide, every claim checked against ground truth with evidence. Grounded facts and context survive resets. Reverse engineering is the proving ground. MCP server + CLI.
[JMLR 2026] "UQLM: A Python Package for Uncertainty Quantification in Large Language Models"
Loki: Open-source solution designed to automate the process of verifying factuality
Awesome-LLM-Robustness: a curated list of Uncertainty, Reliability and Robustness in Large Language Models
Dingo: A Comprehensive AI Data, Model and Application Quality Evaluation Tool
✨✨Woodpecker: Hallucination Correction for Multimodal Large Language Models
RefChecker provides automatic checking pipeline and benchmark dataset for detecting fine-grained hallucinations generated by Large Language Models.
[CVPR'24] HallusionBench: You See What You Think? Or You Think What You See? An Image-Context Reasoning Benchmark Challenging for GPT-4V(ision), LLaVA-1.5, and Other Multi-modality Models
up-to-date curated list of state-of-the-art Large vision language models hallucinations research work, papers & resources
[ICLR'24] Mitigating Hallucination in Large Multi-Modal Models via Robust Instruction Tuning
[ICLR 2025] LLaVA-MoD: Making LLaVA Tiny via MoE-Knowledge Distillation
[ACL 2024] User-friendly evaluation framework: Eval Suite & Benchmarks: UHGEval, HaluEval, HalluQA, etc.
[NeurIPS 2024] Knowledge Circuits in Pretrained Transformers
Explore concepts like Self-Correct, Self-Refine, Self-Improve, Self-Contradict, Self-Play, and Self-Knowledge, alongside o1-like reasoning elevation🍓 and hallucination alleviation🍄.
HaluMem is the first operation level hallucination evaluation benchmark tailored to agent memory systems.
😎 curated list of awesome LMM hallucinations papers, methods & resources.
[ICLR 2025] MLLM can see? Dynamic Correction Decoding for Hallucination Mitigation
To associate your repository with the hallucination topic, visit your repo's landing page and select "manage topics."