惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

Jina AI
Jina AI
博客园 - 【当耐特】
量子位
C
Check Point Blog
博客园 - 叶小钗
博客园 - 聂微东
博客园 - 三生石上(FineUI控件)
让小产品的独立变现更简单 - ezindie.com
让小产品的独立变现更简单 - ezindie.com
Hugging Face - Blog
Hugging Face - Blog
美团技术团队
The Cloudflare Blog
T
Tailwind CSS Blog
人人都是产品经理
人人都是产品经理
月光博客
月光博客
V
V2EX
Last Week in AI
Last Week in AI
酷 壳 – CoolShell
酷 壳 – CoolShell
奇客Solidot–传递最新科技情报
奇客Solidot–传递最新科技情报
IT之家
IT之家
大猫的无限游戏
大猫的无限游戏
有赞技术团队
有赞技术团队
Apple Machine Learning Research
Apple Machine Learning Research
S
SegmentFault 最新的问题
freeCodeCamp Programming Tutorials: Python, JavaScript, Git & More

博客园 - 郭新晨

【三年面试五年模拟】AIGC/LLM/AI Agent算法工程师面试秘籍。涵盖AIGC、LLM大模型、AI Agent、具身智能、传统深度学习、自动驾驶、机器学习、计算机视觉、自然语言处理、强化学习、大数据挖掘、世界模型、元宇宙、AGI等AI行业面试笔试干货经验与核心知识 扩散模型教程 AIGC实战——世界模型(World Model) Github 9 个惊艳的开源 NL2SQL 项目 Curated tutorials and resources for Large Language Models, Text2SQL, Text2DSL、Text2API、Text2Vis and more. Text-to-SQL with llama-index How I Turned a 400-Table SQL DB into a Chat-Friendly RAG System with LlamaIndex 人工智能的数学基础 MATLAB 2025b 安装教程 A curated list of awesome voice conversion, projects and communities Automatically generate, translate, and overlay subtitles for any video Automatically generate and overlay subtitles for any video SoftVC VITS Singing Voice Conversion 有手就行!Sovits AI人声模型训练 A curated roadmap based on my 6 years of experience form zero to become a skilled AI Speech Engineer. This roadmap covers everything from fundamentals to cutting-edge AI新宠DocExt:纯本地文档抽取,开源免费还无依赖!你还在为OCR头疼吗? LangExtract万字实战指南:基于LLM文本结构化工具 Get your documents ready for gen AI HMM隐马尔可夫模型的例子、原理、计算和应用 GUI for a Vocal Remover that uses Deep Neural Networks We provide a PyTorch implementation of the paper Voice Separation with an Unknown Number of Multiple Speakers In which, we present a new method for separating a mixed audio sequence 开源语音分离工具大比拼:人声 VS 背景音乐 ⚔️ - 获取干净训练语音 (数据截至 2025年4月17日)!!! VoiceSplit: Targeted Voice Separation by Speaker-Conditioned Mel Spectrogram VoiceFilter-Lite: Streaming Targeted Voice Separation for On-Device Speech Recognition VoiceSplit: Targeted Voice Separation by Speaker-Conditioned Spectrogram A fork to record speaker output with python. PyAudio with PortAudio for Windows | Extended | Loopback | WASAPI | Latest precompiled Version 头条号爬虫案例 今日头条评论爬虫 - 使用Selenium自动化采集头条文章评论的Python工具 LibriheavyMix Open-source datasets and deep learning models for separating sounds
Kimi-Audio, an open-source audio foundation model excelli...
郭新晨 · 2026-04-29 · via 博客园 - 郭新晨

https://github.com/MoonshotAI/Kimi-Audio

posted @ 2026-04-29 20:01  郭新晨  阅读(13)  评论(0)    收藏  举报

刷新页面返回顶部