惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

Google DeepMind News
Google DeepMind News
博客园 - 司徒正美
OSCHINA 社区最新新闻
OSCHINA 社区最新新闻
V
V2EX
博客园_首页
量子位
博客园 - 三生石上(FineUI控件)
钛媒体:引领未来商业与生活新知
钛媒体:引领未来商业与生活新知
腾讯CDC
T
The Blog of Author Tim Ferriss
博客园 - 聂微东
V
Visual Studio Blog
J
Java Code Geeks
宝玉的分享
宝玉的分享
爱范儿
爱范儿
MongoDB | Blog
MongoDB | Blog
D
Docker
大猫的无限游戏
大猫的无限游戏
Y
Y Combinator Blog
H
Help Net Security
罗磊的独立博客
H
Hackread – Cybersecurity News, Data Breaches, AI and More
Cyber Security Advisories - MS-ISAC
Cyber Security Advisories - MS-ISAC
Blog — PlanetScale
Blog — PlanetScale

PhysioNet News

Delays in reviewing applications for credentialed access PhysioNet as a global platform for biomedical research The Open-Source Engine Behind Modern AI in Medicine – MIT Jameel Clinic Bridge2AI-Voice Adult Cohort Audio Dataset Bridge2AI-Voice Pediatric Cohort Audio Dataset George B. Moody PhysioNet Challenge Seeking Applications for Exceptional Candidates for the Director, National Institute of General Medical Sciences (NIGMS), NIH Use of MIMIC Data with Large Language Models and Online Services Roger Mark and George Moody Receive the 2026 IEEE Biomedical Engineering Award Access Restrictions Under DOJ Data Security Program This repository is under review by NIH for potential modification in compliance with U.S. federal Administration directives. A Dataset for Addressing Patient's Information Needs related to Clinical Course of Hospitalization v1.3 The George B. Moody PhysioNet Challenge 2025 has begun An ethically-sourced, diverse voice dataset linked to health information v3.1.0 MIMIC-IV v3.1 is now available on BigQuery Upgrading MIMIC-IV on BigQuery Delays in reviewing applications for credentialed access George B. Moody PhysioNet Challenge Guidelines for creating derived datasets and models Network issues at MIT, impacting the availability of PhysioNet George B. Moody PhysioNet Challenge 2024: Challenge Opening Duke Critical Care Datathon: 13-14 April 2024 CHIL 2024: Submit your paper by Friday, 16 February 7th Annual Conference on Health, Inference, and Learning George B. Moody PhysioNet Challenge 2024: Challenge Opening Open Source and Validated Computational Tools for Physiological Time Series Analysis RECRUITMENT | inside-heart DARPA Triage Challenge: Qualification extended through Nov 27 Triage Challenge | DARPA Call for partners interested in synthetic patient data
SNOMED CT Entity Linking Benchmark
DrivenData · 2023-12-21 · via PhysioNet News

Introducing DrivenData Benchmarks!

DrivenData Benchmarks provide a rigorous baseline for comparing different approaches to the same problem. Submit multiple different models, see how your methods perform, and contribute to a public body of knowledge about what works.

Overview

Much of the world's healthcare data is stored in free-text documents, such as clinical notes taken by doctors. Because this information is unstructured, it can be difficult to analyze and extract meaningful insights. However, by applying a standardized terminology like SNOMED CT, healthcare organizations can convert this free-text data into a structured format that computers can readily analyze, helping stimulating the development of new medicines, treatment pathways, and better patient outcomes.

One way to analyze clinical notes is to identify and label the portions of each note that correspond to specific medical concepts. This process is called entity linking because it involves identifying candidate spans in the unstructured text (the entities) and linking them to a particular concept in a knowledge base of medical terminology.

Clinical entity linking is inherently challenging. Medical notes are often rife with abbreviations (some of them context-dependent) and assumed knowledge. Furthermore, the target knowledge bases can easily include hundreds of thousands of concepts, many of which occur infrequently leading to a “long tail” effect in the distribution of concepts.

Benchmark task

A synthetic example illustrating how a medical note is annotated with concepts from SNOMED CT.

A synthetic example of medical text and labeled concepts. The concepts are highlighted in green, and the indicated concept IDs, names, and categories are shown.

The objective of this benchmark is to compare how well different approaches can structure the unstructured data in clinical notes for meaningful use and analysis, using the SNOMED CT clinical terminology. Participants will train models based on real-world doctor's notes which have been de-identified and annotated with SNOMED CT concepts by medically trained professionals. This is the largest publicly available dataset of labelled clinical notes! By submitting to this benchmark, you can see how well your model performs and compare it against other approaches.

How to submit to the benchmark

  1. Click the "Register!" button in the sidebar to register for the benchmark.
  2. Get familiar with the problem through the problem description.
  3. Get access to the dataset of clinical notes and the training set of annotations by following the data access instructions.
  4. Create and train your own model. Check out the benchmark blog post or the winning solutions of the original competition for a good place to start!
  5. Package your model files with the code to make predictions based on the runtime repository specification on the code submission format page.
  6. Create a model description by clicking on "My Models" in the sidebar followed by "Create new model". Give your model a name and fill out details about your approach in the abstract field. (Note: this only creates a description of your approach. You do not need to share your actual model.)
  7. Click on your model's "Make new code job submission" to submit your code as a zip archive for containerized execution.
  8. Once you have a successful submission, click "Make public" to share your model description with the community and be added to the benchmark leaderboard!

Rules

The benchmark rules are in place to promote fair benchmarking and useful solutions. If you are ever unsure whether your solution meets the benchmark rules, ask the challenge organizers in the forum or send an email to info@drivendata.org.


SNOMED International logo

In partnership with Veratai and PhysioNet

Veratai logo            PhysioNet logo


Image courtesy of SNOMED International