Moving my blog over to substack! My latest thoughts on AI hype. I welcome any thoughts. substack.com/@alexandergi...
$1,000,000,000,000's isn't worth $30/month.
AI is simply a tool that allows me to do what I could already do faster.
substack.com
Alexander Gibson
@alexdgibson.bsky.social
metascientist www.alexdgibson.com
Moving my blog over to substack! My latest thoughts on AI hype. I welcome any thoughts. substack.com/@alexandergi...
$1,000,000,000,000's isn't worth $30/month.
AI is simply a tool that allows me to do what I could already do faster.
substack.com
Moving my blog over to substack! My latest thoughts on AI hype. I welcome any thoughts. substack.com/@alexandergi...
$1,000,000,000,000's isn't worth $30/month.
AI is simply a tool that allows me to do what I could already do faster.
substack.com
A fabricated dataset hosted on Kaggle was recently removed for breaching copyright, months after it was exposed in research led by #AusHSI PhD student @alexdgibson.bsky.social with @aidybarnett.bsky.social, @nicolewhite.bsky.social and Gary Collins. Find out more 👇
The online data repository Kaggle has removed a problematic dataset for violating intellectual property rights months after it was flagged by sleuths for containing celebrity photos and lacking information on data provenance or confirmed medical diagnoses.
I’ve not been excited about technology in a long while. @frame.work modular, upgradable, repairable laptops have got me SO EXCITED. Respecting consumers and the planet will always wins. Putting my MacBook up for sale and making an upgrade!
Yeah, that's not good. Having looked into how a lot of medical guidance etc gets made manually, it doesn't surprise me. But it's not good.
Evidence of unreliable data and poor data provenance in clinical prediction model research and clinical practice - BMC Medicine
Background Clinical prediction models are often created using large routinely collected datasets. It is essential that prediction models are developed with appropriate data and methods and transparent...
link.springer.com
NEW #opensaccess paper "Evidence of unreliable data and poor data provenance in clinical prediction model research and clinical practice" link.springer.com/article/10.1... #researchintegrity #dataprovenance #predictionmodels #predictiveAI
🚨 Protocol -> preregistration -> preprint -> publication 🚨 125 clinical prediction model articles developed on unreliable data on stroke and diabetes. Since preprinting our results 7 articles have been retracted. doi.org/10.1186/s129...
Some of the models seem to have been used in clinical settings although it’s not clear whether this has led to flawed diagnoses go.nature.com/3Q7scE0
Dozens of AI disease-prediction models were trained on dubious data
The models are designed to predict someone’s risk of diabetes or stroke. A few might already have been used on patients.
go.nature.com
Outstanding work by @alexdgibson.bsky.social to uncover the use of highly dubious data sets from @kaggle.com being used in hundreds of research papers and potentially even informing clinical practice. If you're re-using data, take the time to confirm that it's real. www.nature.com/articles/d41....
Dozens of AI disease-prediction models were trained on dubious data
The models are designed to predict someone’s risk of diabetes or stroke. A few might already have been used on patients.
nature.com
Markus Englund built a tool for identifying repetitive sequences in scientific datasets! Read up here: www.sciencedetective.org/scientific-d...
Scientific datasets are riddled with copy-paste errors
Initial results from scanning through Excel files belonging to 600 published scientific papers.
sciencedetective.org
Evidence of unreliable data and poor data provenance in published clinical prediction models and clinical practice 📣 My first PhD study is available on MedRxiv: www.medrxiv.org/content/10.6...
Evidence of Unreliable Data and Poor Data Provenance in Clinical Prediction Model Research and Clinical Practice
Clinical prediction models are often created using large routinely collected datasets. It is essential that prediction models are developed with appropriate data and methods and transparently reported...
medrxiv.org
Evidence of unreliable data and poor data provenance in published clinical prediction models and clinical practice 📣 My first PhD study is available on MedRxiv: www.medrxiv.org/content/10.6...
Evidence of Unreliable Data and Poor Data Provenance in Clinical Prediction Model Research and Clinical Practice
Clinical prediction models are often created using large routinely collected datasets. It is essential that prediction models are developed with appropriate data and methods and transparently reported...
medrxiv.org
New Blog: Learning R for Good Research Practices‼️ Read Part 1: alexdgibson.com/blog/r_1/ This is the first of a multipart series where I go into my experiences learning R, highlighting tips and resources that were essential in my learning.
Learning R for Good Research Practices: Part 1 | Alexander Gibson
In this multi-part blog series I outline some steps to help start learning R. Part 1 is a background on R, packages and resources and tips on how to get started learning.
alexdgibson.com
We need more people reading more books! Such an important skill and has done so much for me. There’s nothing better than a good book, a change in perspective, beliefs or the generation of new ideas.
Last week, #AusHSI PhD student @alexdgibson.bsky.social presented at #AIMOS2025, highlighting a new global issue of unreliable data within #clinicalprediction model research and clinical practice, which has potential to influence patient outcomes and evidence-based decisions. @aimosinc.bsky.social
First up: Alexander Gibson @alexdgibson.bsky.social: Poor Data Provenance in Published Clinical Prediction Model Research. I was seriously concerned about some of the Kaggle Datasets - those on stroke and diabetes raised concerns about data being fake. #AIMOS2025
📣 New Blog 📣 Learn more about my #PhD Research and work @aushsi.bsky.social www.aushsi.org.au/a-model-of-g...
A model of good research practice in clinical prediction - AusHSI
When it comes to health and medical research, doing the right thing is critical, especially when it impacts patient outcomes. Alexander Gibson's research focuses on identifying statistical and researc...
aushsi.org.au
Ok, time for a short thread about this paper. My sense over the past six months or so is that chain-of-thought prompting as used in e.g. ChatGPT o.3 improves substantially upon previous systems such as ChatGPT 4.o, at least for certain tasks. But how revolutionary is it?
If I have time I'll put together a more detailed thread tomorrow, but for now, I think this new paper about limitations of Chain-of-Thought models could be quite important. Worth a look if you're interested in these sorts of things. ml-site.cdn-apple.com/papers/the-i...
🚨Job Alert! 🚨 #AusHSI is seeking a new Research Project Officer to work collaboratively with a team of leading #healthservices researchers and partners, providing support across a range of major research projects to deliver exceptional outcomes. 👉 Find out more and apply today: bit.ly/4k1ba4h
Reading over my brother’s undergraduate assignment. Criteria has a section: “Is there an appropriate amount of detail in this section so that you could replicate this study?” I don’t remember learning about replication in my undergrad! Nice to see it getting some air time.
If you're currently working—or have worked in the past 5 years—in consulting or collaborative research as a biostatistician, we’d love to hear from you. 📅 Closes 11th July This survey has been approved by the QUT Human Research Ethics Committee (approval #9691).
International #ResearchIntegrity conference taking place in beautiful Sydney Australia on 16-18 November 2025 🐨 🌏 Great speakers including @elisabethbik.bsky.social @jdwilko.bsky.social @jasonchin.bsky.social For more info contact @simongandevia.bsky.social 🧪
International Research Integrity Conference. I am arranging this in Sydney in November 16-18th 2025. researchintegrityconf.com Note the excellent speaker list. Contact me!
Lazy cross post, please help. I have a bad feeling about this.
I thought my first post on Bluesky should be something positive and motivational
I’ve not been on Bluesky for a while but there seems to be so many interesting and engaging people! Lots of good #research
Should my PhD Exist? My latest blog ✍️: alexdgibson.com/blog/good_sc... “My hope is one day science will be so rigorous, that all focus can be on progressing science forward, not identifying common problems.” #researchintegrity #metaresesrch #phd
Should My PhD Exist? | Alexander Gibson
We don’t need more science, we need better science.
alexdgibson.com
#AusHSI Prof Will Parsonage is one of more than 20 experts from across the world involved in a new @thelancet.bsky.social Commission calling on the medical profession to treat coronary #heartdisease as a lifelong condition, which has the potential to save 8.7 million lives every year: bit.ly/3XEMd5G
*NEW PAPER* PROBAST+AI: an updated quality, risk of bias & applicability assessment tool for prediction models using regression or AI methods PROBAST+AI consists of two distinct parts: - model development (quality assessment tool) - model evaluation (risk of bias tool) www.bmj.com/content/388/...
PROBAST+AI: an updated quality, risk of bias, and applicability assessment tool for prediction models using regression or artificial intelligence methods
The Prediction model Risk Of Bias ASsessment Tool (PROBAST) is used to assess the quality, risk of bias, and applicability of prediction models or algorithms and of prediction model/algorithm studies....
bmj.com