HN user

jaan

85 karma

Machine learning for health & science.

https://jaan.io

https://twitter.com/thejaan

Posts14
Comments23
View on HN
Googlebook 2 months ago

Wow, I’m in the same boat - do you mind sharing more about how you did it? I was thinking about that too (I’m 197cm) and would love to learn!

Open to any feedback on this here or over email (jaan@onefact.org)!

I quit academia to start a non-profit focused on using open source to analyze the public hospital price transparency data.

We are now making similar dashboards for every hospital in the country, and need all the help we can get if you would be interested in using the latest geospatial mapping tools, databases (duckdb) and large language models to make sense of this massive amount of data.

Through a data bounty (https://www.dolthub.com/repositories/onefact/paylesshealth) we collected 4000+ hospital price sheets and made them public here: https://data.payless.health/#hospital_price_transparency/. This was on HN previously.

Grateful for all your support so far!!

Yes! We are working on this and integrating with the OMOP common data model, to be able to link the health outcomes in our data partners' clinical repositories to the cost of care. For example, we work with the NIH All of Us study for outcome data (joinallofus.org -- I signed up both to contribute to this science and to get my whole genome sequenced free!)

If you look at the files, many of them are not compliant, and so we need to figure out what the associated line item corresponds to: a CPT code? HCPCS code? ICD code? etc :)

Here's an example NLP tool I helped build we're using to do this: https://arxiv.org/abs/1904.05342 -- it's in several pipelines now for data annotation and crowdsourcing.

I'm surprised no one has mentioned Cadence & Slang by Nick Disabato: http://cadence.cc/

I found it a gentle but thorough introduction with great references for further reading. I wish I had read this before trying to design apps for the first time - it would have saved a lot of headache.

Thanks – I agree with your worries about misinterpretation, especially in regard to flawed study design. We're not trying to be the be-all end-all, but hopefully a decent starting point for further research into the subtleties of a specific topic. That's in addition to trying to make the science more accessible by having it appeal to a wider audience.

We're still narrowing down our content guidelines, so would love your input! Feel free to ping me.