Jove
Visualize
お問い合わせ
JoVE
x logofacebook logolinkedin logoyoutube logo
JoVEについて
概要リーダーシップブログJoVEヘルプセンター
著者向け
出版プロセス編集委員会範囲と方針査読よくある質問投稿
図書館員向け
推薦の声購読アクセスリソース図書館諮問委員会よくある質問
研究
JoVE JournalMethods CollectionsJoVE Encyclopedia of Experimentsアーカイブ
教育
JoVE CoreJoVE BusinessJoVE Science EducationJoVE Lab Manual教員リソースセンター教員サイト
利用規約
プライバシーポリシー
ポリシー

関連する概念動画

Qualitative Analysis01:10

Qualitative Analysis

674
Qualitative analysis is the process of identifying elements, ions, or compounds in an unknown sample. It is the first and most fundamental type of analysis based on the hierarchy of analytical goals. This hierarchy is significant as it provides a structured approach to scientific research, with qualitative analysis serving as the initial step, providing essential information before moving on to quantitative or other forms of analysis.
There are two main approaches to qualitative analysis:...
674

こちらも読む

関連記事

共著者、ジャーナル、引用グラフによってこの研究に関連する記事。

並び替え
Same author

A S.C.O.R.E. framework for evaluating open-ended responses from large language models in healthcare.

Cell reports. Medicine·2026
Same author

Regional and temporal trends in antimicrobial susceptibility among isolates from bacterial keratitis: a systematic review and meta-analysis.

The Lancet. Microbe·2026
Same author

Corrigendum to "Oculomics and AI: The eye as a biomarker for health span" [Asia-Pac J Ophthalmol 15 (1) (2026) 100282].

Asia-Pacific journal of ophthalmology (Philadelphia, Pa.)·2026
Same author

AI-induced never-skilling in medical education.

Nature medicine·2026
Same author

How to meaningfully evaluate AI in clinical medicine.

Nature medicine·2026
Same author

Application of artificial intelligence and robotics in ophthalmic practice.

Asia-Pacific journal of ophthalmology (Philadelphia, Pa.)·2026

関連する実験動画

Updated: Sep 9, 2025

Augmenting Large Language Models via Vector Embeddings to Improve Domain-Specific Responsiveness
03:14

Augmenting Large Language Models via Vector Embeddings to Improve Domain-Specific Responsiveness

Published on: December 6, 2024

680

眼科の大型言語モデルの評価:定量的な対質的な方法

Ting Fang Tan1, Arun J Thirunavukarasu2, Chrystie Quek1

  • 1Singapore National Eye Centre, Singapore Eye Research Institute, Singapore, Singapore.

Current opinion in ophthalmology
|September 5, 2025
PubMed
まとめ

眼科における大型言語モデル (LLM) と生成型人工知能 (AI) の評価には,精度を超えた多様な指標が必要です. 標準化されたベンチマークと厳選されたデータセットは 堅実な臨床検証と統合に不可欠です

キーワード:
GPT について評価する人工知能大型言語モデル眼科について

さらに関連する動画

A Method to Quantify Visual Information Processing in Children Using Eye Tracking
09:47

A Method to Quantify Visual Information Processing in Children Using Eye Tracking

Published on: July 9, 2016

17.6K
Author Spotlight: Quantifying Rough Eye Phenotypes in Drosophila Models of Amyotrophic Lateral Sclerosis with Frontotemporal Dementia
05:25

Author Spotlight: Quantifying Rough Eye Phenotypes in Drosophila Models of Amyotrophic Lateral Sclerosis with Frontotemporal Dementia

Published on: October 4, 2024

1.0K

関連する実験動画

Last Updated: Sep 9, 2025

Augmenting Large Language Models via Vector Embeddings to Improve Domain-Specific Responsiveness
03:14

Augmenting Large Language Models via Vector Embeddings to Improve Domain-Specific Responsiveness

Published on: December 6, 2024

680
A Method to Quantify Visual Information Processing in Children Using Eye Tracking
09:47

A Method to Quantify Visual Information Processing in Children Using Eye Tracking

Published on: July 9, 2016

17.6K
Author Spotlight: Quantifying Rough Eye Phenotypes in Drosophila Models of Amyotrophic Lateral Sclerosis with Frontotemporal Dementia
05:25

Author Spotlight: Quantifying Rough Eye Phenotypes in Drosophila Models of Amyotrophic Lateral Sclerosis with Frontotemporal Dementia

Published on: October 4, 2024

1.0K

科学分野:

  • 眼科について
  • 人工知能
  • 医療情報学

背景:

  • 大型言語モデル (LLM) と生成型人工知能 (AI) は,臨床眼科においてますます適用されています.
  • これらのAIツールのパフォーマンスを評価することは,安全で効果的な医療への統合に不可欠です.
  • 現在の評価方法は,AIの性能の他の重要な側面を無視して,しばしば正確性だけに焦点を当てています.

研究 の 目的:

  • 眼科におけるLLMおよび生成AIアプリケーションの評価メトリックの重要性を検討し,強調する.
  • 一般的に採用されている定量・定性評価指標について議論する.
  • この分野における強力なAI評価を妨げている課題を特定する.

主な方法:

  • 眼科におけるLLMとAI評価に関する既存の文献の体系的なレビュー.
  • 出版された研究で用いられた定量・定性指標の分析
  • 現在の評価慣行における共通の課題と限界を特定する.

主要な成果:

  • ジェネラティブAIは様々な眼科の臨床応用において有望なパフォーマンスを示しています
  • 精度を超えて,定量と質のメトリックは,LLMの出力をより包括的に評価します.
  • 主な課題は,標準化された基準の欠如と,高品質の臨床データセットの限られた可用性です.

結論:

  • 評価メトリックのスペクトルは存在するが,標準化は欠けている.
  • データセットの管理とベンチマーク開発の課題に取り組むことが不可欠です.
  • 臨床的なAIアプリケーションを検証し,眼科の実践への統合を容易にするには,堅実で領域特有の評価が不可欠です.