Skip to content

fix(feishu): 语音消息映射为 VOICE 类型以启用 STT 转录 - #30174

Closed
wangzengzhi wants to merge 1 commit into
NousResearch:mainfrom
wangzengzhi:fix/feishu-voice-stt-mapping
Closed

fix(feishu): 语音消息映射为 VOICE 类型以启用 STT 转录#30174
wangzengzhi wants to merge 1 commit into
NousResearch:mainfrom
wangzengzhi:fix/feishu-voice-stt-mapping

Conversation

@wangzengzhi

Copy link
Copy Markdown

飞书将语音消息标记为 AUDIO 类型,但 Gateway 运行管线中 AUDIO 被当作文件附件处理(不走 STT 转写管道)。 修改 _resolve_media_message_type 将 audio/* 媒体映射为 VOICE 类型, 使飞书语音消息能正确进入 _enrich_message_with_transcription 转写流程。

What does this PR do?

Related Issue

Fixes #

Type of Change

  • 🐛 Bug fix (non-breaking change that fixes an issue)
  • ✨ New feature (non-breaking change that adds functionality)
  • 🔒 Security fix
  • 📝 Documentation update
  • ✅ Tests (adding or improving test coverage)
  • ♻️ Refactor (no behavior change)
  • 🎯 New skill (bundled or hub)

Changes Made

How to Test

Checklist

Code

  • I've read the Contributing Guide
  • My commit messages follow Conventional Commits (fix(scope):, feat(scope):, etc.)
  • I searched for existing PRs to make sure this isn't a duplicate
  • My PR contains only changes related to this fix/feature (no unrelated commits)
  • I've run pytest tests/ -q and all tests pass
  • I've added tests for my changes (required for bug fixes, strongly encouraged for features)
  • I've tested on my platform:

Documentation & Housekeeping

  • I've updated relevant documentation (README, docs/, docstrings) — or N/A
  • I've updated cli-config.yaml.example if I added/changed config keys — or N/A
  • I've updated CONTRIBUTING.md or AGENTS.md if I changed architecture or workflows — or N/A
  • I've considered cross-platform impact (Windows, macOS) per the compatibility guide — or N/A
  • I've updated tool descriptions/schemas if I changed tool behavior — or N/A

For New Skills

  • This skill is broadly useful to most users (if bundled) — see Contributing Guide
  • SKILL.md follows the standard format (frontmatter, trigger conditions, steps, pitfalls)
  • No external dependencies that aren't already available (prefer stdlib, curl, existing Hermes tools)
  • I've tested the skill end-to-end: hermes --toolsets skills -q "Use the X skill to do Y"

Screenshots / Logs

飞书将语音消息标记为 AUDIO 类型,但 Gateway 运行管线中 AUDIO 被当作文件附件处理(不走 STT 转写管道)。
修改 _resolve_media_message_type 将 audio/* 媒体映射为 VOICE 类型,
使飞书语音消息能正确进入 _enrich_message_with_transcription 转写流程。
@alt-glitch alt-glitch added type/bug Something isn't working platform/feishu Feishu / Lark adapter tool/tts Text-to-speech and transcription comp/gateway Gateway runner, session dispatch, delivery P2 Medium — degraded but workaround exists labels May 22, 2026
@alt-glitch

Copy link
Copy Markdown
Collaborator

Duplicate of #29235 (and #29295). Same fix — mapping Feishu audio to VOICE for STT pipeline.

Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Labels

comp/gateway Gateway runner, session dispatch, delivery P2 Medium — degraded but workaround exists platform/feishu Feishu / Lark adapter tool/tts Text-to-speech and transcription type/bug Something isn't working

Projects

None yet

Development

Successfully merging this pull request may close these issues.

2 participants