Working in industrial research is usually very motivating but occasionally it is also frustrating. You’ve just done something really cool but you’re not allowed to tell anybody outside the company about it. Indeed, in a small company there might not be anybody inside of the company who can even appreciate it!
I have worked on roughly 4 really cool projects since leaving academia at the end of 2017. And apart from some basic mentions in my blog (e.g. here and here) most of what I have done has been known only to a few key stakeholders.
I had the opportunity to talk recently with a relatively advanced researcher in machine learning methods. The conversation turned briefly to the study of embeddings when he mentioned that most of his work involves things that can be embedded in Euclidean space. Since I’ve been spending a bit of time thinking about embeddings recently, I asked him some questions to get the official ML take on the subject. I was resonably gratified to learn that – although most ML engineers don’t think much about embeddings – the research on this topic considers the embedding to be tightly bound to the network architecture. It is not possible to study abstract embeddings, divorced from applications. I fully agree with this point-of-view.
Randomised controlled trials (RCTs) have been the gold standard for statistical evidence, of treatment effect, for over 100 years. Their strength is in their attempt to avoid major sources of bias in a comparison of the evidence. However, they are costly to run, particularly in the domain of personalised medicine, to which medical AI products typically belong.
I have a short thought, stemming from a combination of projects that I’m working on at the moment, and I want to share it.
The current trend towards Causality in AI is very attractive to people like me. It matches our personal biases and views of the world. However, it is lacking a natural heuristic. How do we decide how much resources to devote to alternative models of the world, as we gather evidence as to their accuracy?
Like I say, I have a number of parallel projects, many of which address exactly this question on technical and biological levels.
There is something from the world of business, studying entrepreneurship, which might be a better heuristic than any normative model I can come up with. Effectual entrepreneurship is a perspective on entrepreneurship, studying highly successful repeat entrepreneurs (eg. Elon Musk), which establishes control, rather than planning, at the core of entrepreneurial activities.
First a mea culpa, I have a huge backlog of relatively heavy articles that I really want to add to the blog. But I’ve been busy getting married – congratulations to me – and I didn’t have enough time. I strongly believe in following relatively strict guidelines on writing and editing articles, where I set myself deadlines and avoid over-writing on topics – it is just a blog after all – but for deep insights I do also have a minimum standard that I want to be able to produce before I’m willing to hit the Publish button.
I am beginning a new project this week, the topic is Causal Inference. This is something I have been reading about, and wrestling with, for quite some time. Now seems a good point to take some time out, form a project, and see what I can get done on the topic.
How do I really feel about this topic? I think that I can only work out the answer to this question by writing about it.
My suspicion is that those who shout loudest about personalised medicine know least about it. I fear that the promises being made publicly are categorically not possible. My hope is that I am wrong on this.