Mitigating medical bias in large language models by prompt engineering: an empirical study of effectiveness and trade-offs
{{output}}
Large language models (LLMs) demonstrate expert-level performance in various medical scenarios, yet their outputs can exhibit bias against groups or individuals with specific sensitive attributes, posing risks to patient safety and undermining trust in LLMs fo... ...