back
user profile
apwheele
1,887karma·374submissions·November 10, 2021
about
Data scientist. Former academic in criminal justice field.
Personal blog at https://andrewpwheeler.com/
Consulting at https://crimede-coder.com/
Large Language Models for Mortals: A Practical Guide for Analysts (book), https://crimede-coder.com/blogposts/2026/LLMsForMortals
recent activity (374 total)
comment
For a newer list I would add the mapbox API as well. So I work in data analytics, not so much web-mapping. For those applications, IMO local solutions, like ESRI, are good options if you are limited t…
comment
Care to share the contexts in which someone needs a zero-shot model for time series? I have just never come across one in which you don't have some historical data to fit a model and go from ther…
comment
I feel the same way, and a good follow up book about more general communication is Trees, maps and theorems by Doumont, https://andrewpwheeler.com/2016/12/05/review-of-…
comment
IMO if doing this, you should avoid text in the charts entirely (as the title can sometimes I think lead the models astray, such as the clustering title I think will bias it to find clusters even if n…
comment
Wayback url, https://web.archive.org/web/20241118131800/https://www.newso... …
comment
My companies looks similarish to the recent screenshot, but it is a hellscape of a billion options and poor search functionality. To the extent I just need to ask a person the right link or tree searc…
comment
A tell for fake firms in my local newspaper is they ask for a snail mail resume. These appear to me to be more like shell companies submitting multiple H1Bs as far as I can tell though, not legit firm…
comment
I agree understanding KM is a very good place to start survival analysis. Many examples in my business I have for KM the censoring is due to certain events taking along time (auditing healthcare claim…
comment
Care to elaborate on this? So this post does not save the resulting weight, so you don't use that in any subsequent calculations. You would just treat the result as a simple random sample. So it …
comment
Very nice, another pro-tip for folks is that you can set the weights to get approximate stratified sampling. So say group A had 100,000 rows, and group B had 10,000 rows, and you wanted each in the re…
comment
I'm sorry but this model is wild and highly misleading. Based on one variable and 7 observations, days since you have started, you have fit a model with `2 + 128 + 64^2` parameters. It happens th…
comment
So have not dealt with McKinsey specifically, but other MBA/Consultant types I feel like they treat software engineering like we are ditch diggers -- they just request things and expect it to mag…
comment
So the message on my desktop is below where I can see, and ditto when I preview in phone (either orientation). So still a problem. IMO I would just make the input box smaller. And maybe do a legend fo…