• About
  • Advertise
  • Privacy & Policy
  • Contact
HK Businesswire
  • Home
  • News
    • All
    • Business
    • Politics
    • PR Newswire
    • Science
    • World
    Postal summit held

    Postal summit held

    RemeGen’s Telitacicept (RC18) Received Orphan Drug Designation from EMA for Myasthenia Gravis

    RemeGen’s Telitacicept (RC18) Received Orphan Drug Designation from EMA for Myasthenia Gravis

    Hexagon launches AEON, a humanoid built for industry

    Hexagon launches AEON, a humanoid built for industry

    RuggON Unveils 12-inch SOL 7: The World’s First Rugged Tablet Powered by Intel® Arrow Lake Processors

    RuggON Unveils 12-inch SOL 7: The World’s First Rugged Tablet Powered by Intel® Arrow Lake Processors

    HK gains 9 Martin Barnes Awards

    HK gains 9 Martin Barnes Awards

    Polus holds €425 million initial close for third CLO equity fund

    Polus holds €425 million initial close for third CLO equity fund

    Trending Tags

    • Trump Inauguration
    • United Stated
    • White House
    • Market Stories
    • Election Results
  • PR Newswire
  • Business
  • World
  • Entertainment
  • Sports
  • Tech
    • All
    • Apps
    • Gadget
    • Mobile
    • Startup

    Deloitte: Over 40% of Family Offices Prioritise Tech Amid Digital Transformation

    PwC: AI-Exposed Jobs See Surge in Demand, Pay, and Productivity

    PwC: AI-Exposed Jobs See Surge in Demand, Pay, and Productivity

    Hong Kong Student Criticised for Using Outsourced AI Project to Win STEM Awards

    Xiaomi SU7 Ultra Becomes Fastest Mass-Produced EV on Nürburgring Nordschleife

    MPF at 25: PwC and HKRSA Urge Bold Reform for Hong Kong’s Retirement System

    CrowdStrike Shares Dip Despite Strong Q1 Earnings Amid Soft Revenue Guidance

    Trending Tags

    • Nintendo Switch
    • CES 2017
    • Playstation 4 Pro
    • Mark Zuckerberg
  • Feature
No Result
View All Result
  • Home
  • News
    • All
    • Business
    • Politics
    • PR Newswire
    • Science
    • World
    Postal summit held

    Postal summit held

    RemeGen’s Telitacicept (RC18) Received Orphan Drug Designation from EMA for Myasthenia Gravis

    RemeGen’s Telitacicept (RC18) Received Orphan Drug Designation from EMA for Myasthenia Gravis

    Hexagon launches AEON, a humanoid built for industry

    Hexagon launches AEON, a humanoid built for industry

    RuggON Unveils 12-inch SOL 7: The World’s First Rugged Tablet Powered by Intel® Arrow Lake Processors

    RuggON Unveils 12-inch SOL 7: The World’s First Rugged Tablet Powered by Intel® Arrow Lake Processors

    HK gains 9 Martin Barnes Awards

    HK gains 9 Martin Barnes Awards

    Polus holds €425 million initial close for third CLO equity fund

    Polus holds €425 million initial close for third CLO equity fund

    Trending Tags

    • Trump Inauguration
    • United Stated
    • White House
    • Market Stories
    • Election Results
  • PR Newswire
  • Business
  • World
  • Entertainment
  • Sports
  • Tech
    • All
    • Apps
    • Gadget
    • Mobile
    • Startup

    Deloitte: Over 40% of Family Offices Prioritise Tech Amid Digital Transformation

    PwC: AI-Exposed Jobs See Surge in Demand, Pay, and Productivity

    PwC: AI-Exposed Jobs See Surge in Demand, Pay, and Productivity

    Hong Kong Student Criticised for Using Outsourced AI Project to Win STEM Awards

    Xiaomi SU7 Ultra Becomes Fastest Mass-Produced EV on Nürburgring Nordschleife

    MPF at 25: PwC and HKRSA Urge Bold Reform for Hong Kong’s Retirement System

    CrowdStrike Shares Dip Despite Strong Q1 Earnings Amid Soft Revenue Guidance

    Trending Tags

    • Nintendo Switch
    • CES 2017
    • Playstation 4 Pro
    • Mark Zuckerberg
  • Feature
No Result
View All Result
HK Businesswire
No Result
View All Result
Home News Science

Making AI models more trustworthy for high-stakes settings

David Lee by David Lee
1 May 2025
in Science
0
Making AI models more trustworthy for high-stakes settings
0
SHARES
2
VIEWS
Share on FacebookShare on Twitter

The ambiguity in medical imaging can present major challenges for clinicians who are trying to identify disease. For instance, in a chest X-ray, pleural effusion, an abnormal buildup of fluid in the lungs, can look very much like pulmonary infiltrates, which are accumulations of pus or blood.An artificial intelligence model could assist the clinician in X-ray analysis by helping to identify subtle details and boosting the efficiency of the diagnosis process. But because so many possible conditions could be present in one image, the clinician would likely want to consider a set of possibilities, rather than only having one AI prediction to evaluate.One promising way to produce a set of possibilities, called conformal classification, is convenient because it can be readily implemented on top of an existing machine-learning model. However, it can produce sets that are impractically large. MIT researchers have now developed a simple and effective improvement that can reduce the size of prediction sets by up to 30 percent while also making predictions more reliable.Having a smaller prediction set may help a clinician zero in on the right diagnosis more efficiently, which could improve and streamline treatment for patients. This method could be useful across a range of classification tasks — say, for identifying the species of an animal in an image from a wildlife park — as it provides a smaller but more accurate set of options.“With fewer classes to consider, the sets of predictions are naturally more informative in that you are choosing between fewer options. In a sense, you are not really sacrificing anything in terms of accuracy for something that is more informative,” says Divya Shanmugam PhD ’24, a postdoc at Cornell Tech who conducted this research while she was an MIT graduate student.Shanmugam is joined on the paper by Helen Lu ’24; Swami Sankaranarayanan, a former MIT postdoc who is now a research scientist at Lilia Biosciences; and senior author John Guttag, the Dugald C. Jackson Professor of Computer Science and Electrical Engineering at MIT and a member of the MIT Computer Science and Artificial Intelligence Laboratory (CSAIL). The research will be presented at the Conference on Computer Vision and Pattern Recognition in June.Prediction guaranteesAI assistants deployed for high-stakes tasks, like classifying diseases in medical images, are typically designed to produce a probability score along with each prediction so a user can gauge the model’s confidence. For instance, a model might predict that there is a 20 percent chance an image corresponds to a particular diagnosis, like pleurisy.But it is difficult to trust a model’s predicted confidence because much prior research has shown that these probabilities can be inaccurate. With conformal classification, the model’s prediction is replaced by a set of the most probable diagnoses along with a guarantee that the correct diagnosis is somewhere in the set.But the inherent uncertainty in AI predictions often causes the model to output sets that are far too large to be useful.For instance, if a model is classifying an animal in an image as one of 10,000 potential species, it might output a set of 200 predictions so it can offer a strong guarantee.“That is quite a few classes for someone to sift through to figure out what the right class is,” Shanmugam says.The technique can also be unreliable because tiny changes to inputs, like slightly rotating an image, can yield entirely different sets of predictions.To make conformal classification more useful, the researchers applied a technique developed to improve the accuracy of computer vision models called test-time augmentation (TTA).TTA creates multiple augmentations of a single image in a dataset, perhaps by cropping the image, flipping it, zooming in, etc. Then it applies a computer vision model to each version of the same image and aggregates its predictions.“In this way, you get multiple predictions from a single example. Aggregating predictions in this way improves predictions in terms of accuracy and robustness,” Shanmugam explains.Maximizing accuracyTo apply TTA, the researchers hold out some labeled image data used for the conformal classification process. They learn to aggregate the augmentations on these held-out data, automatically augmenting the images in a way that maximizes the accuracy of the underlying model’s predictions.Then they run conformal classification on the model’s new, TTA-transformed predictions. The conformal classifier outputs a smaller set of probable predictions for the same confidence guarantee.“Combining test-time augmentation with conformal prediction is simple to implement, effective in practice, and requires no model retraining,” Shanmugam says.Compared to prior work in conformal prediction across several standard image classification benchmarks, their TTA-augmented method reduced prediction set sizes across experiments, from 10 to 30 percent.Importantly, the technique achieves this reduction in prediction set size while maintaining the probability guarantee.The researchers also found that, even though they are sacrificing some labeled data that would normally be used for the conformal classification procedure, TTA boosts accuracy enough to outweigh the cost of losing those data.“It raises interesting questions about how we used labeled data after model training. The allocation of labeled data between different post-training steps is an important direction for future work,” Shanmugam says.In the future, the researchers want to validate the effectiveness of such an approach in the context of models that classify text instead of images. To further improve the work, the researchers are also considering ways to reduce the amount of computation required for TTA.This research is funded, in part, by the Wistrom Corporation.

Tags: Science
David Lee

David Lee

Read More

Closing in on superconducting semiconductors

Closing in on superconducting semiconductors

17 June 2025
A brief history of the global economy, through the lens of a single barge

A brief history of the global economy, through the lens of a single barge

17 June 2025
  • Trending
  • Comments
  • Latest

Hong Kong Student Criticised for Using Outsourced AI Project to Win STEM Awards

16 June 2025

Macau Enforces 183-Day Residency Rule for 2025 Wealth Partaking Scheme

29 May 2025

Xiaomi SU7 Ultra Becomes Fastest Mass-Produced EV on Nürburgring Nordschleife

11 June 2025
Over 150 firms hoping to list in Hong Kong: HKEX

Over 150 firms hoping to list in Hong Kong: HKEX

28 May 2025
Postal summit held

Postal summit held

17 June 2025
RemeGen’s Telitacicept (RC18) Received Orphan Drug Designation from EMA for Myasthenia Gravis

RemeGen’s Telitacicept (RC18) Received Orphan Drug Designation from EMA for Myasthenia Gravis

17 June 2025
Hexagon launches AEON, a humanoid built for industry

Hexagon launches AEON, a humanoid built for industry

17 June 2025
RuggON Unveils 12-inch SOL 7: The World’s First Rugged Tablet Powered by Intel® Arrow Lake Processors

RuggON Unveils 12-inch SOL 7: The World’s First Rugged Tablet Powered by Intel® Arrow Lake Processors

17 June 2025

Recent News

Postal summit held

Postal summit held

17 June 2025
RemeGen’s Telitacicept (RC18) Received Orphan Drug Designation from EMA for Myasthenia Gravis

RemeGen’s Telitacicept (RC18) Received Orphan Drug Designation from EMA for Myasthenia Gravis

17 June 2025
Hexagon launches AEON, a humanoid built for industry

Hexagon launches AEON, a humanoid built for industry

17 June 2025
RuggON Unveils 12-inch SOL 7: The World’s First Rugged Tablet Powered by Intel® Arrow Lake Processors

RuggON Unveils 12-inch SOL 7: The World’s First Rugged Tablet Powered by Intel® Arrow Lake Processors

17 June 2025
HK Businesswire

Stay ahead with the latest insights on Hong Kong’s economy, finance, and investments. From market trends to policy updates, we bring you in-depth analysis and expert opinions.

📩 Subscribe to our newsletter for exclusive updates.
📍 Follow us on social media for real-time news.
📧 Contact us: info@hongkong-invest.com

Follow Us

  • About
  • Advertise
  • Privacy & Policy
  • Contact

© 2025 by HKBusinesswire.com

No Result
View All Result

© 2025 by HKBusinesswire.com