why aren't all 10 resolved? a statement only gets an assessment when the public
record can support or contradict it. opinions and what-ifs never can, and 0 checkable
ones are still open, waiting for their date. predictions held up or didn't;
assertions are supported or contradicted. on every card:
▮▮▮▮▮ certainty ·
▮▮▮▮▮ debate potential. speakers are clickable
Disclosure
Bloomberg chose Apache Solr over Elasticsearch to maintain open-source control
“One was, we were looking for something that was truly open source, where we had a lot more control over open source and, sorry, source code, and a lot more saying how the direction goes.”
Assertion Not checkable as stated
Bloomberg cannot rely on large-scale A/B testing or usage data
“So a lot of the luxury that other companies have, like doing large scale machine learning based on massive amount of usage data, or doing large scale A-B testing to decide which model works better, doesn't work for us.”
Disclosure
Learning-to-rank algorithms continuously improve search models without manual human tweaking
“We had a huge improvement once we started using it, but more than that improvement, the good part about learning to rank is it always continues to improve your model without constant need of people tweaking it.”
Insight
Embedding machine learning reranking inside Apache Solr improves search performance
“And also, it reduces the hop between two services, so your performance gets much better.”
Disclosure
Bloomberg limits unsupervised learning to data bootstrapping due to precision needs
“We use unsupervised learning a lot to Sort of bootstrap our data. For example, word to whack, right? It's a perfect unsupervised learning algorithm. We use that a lot to, for query reformulation and all that, but, ah, it's a little risky for something as high …”
Disclosure
Bloomberg plans to consolidate terminal search into a single ranked list
“And we want to get to a place where you can search all that in one screen, and we will give you that information in one single ranked list, so you don't have to learn more and more different places to search for data.”
Assertion Supported
Bloomberg contributed code to the last fifteen Apache Solr releases
“I think about the last 15, 15 or 16 versions of Solr have had some sort of our code in it”
Disclosure
Bloomberg uses LambdaMART decision tree algorithms for terminal search ranking
“The one we use is called Lambda Mart, which is based on decision trees, or gradient, regression trees, actually.”
Disclosure
Bloomberg uses interleaving instead of A/B testing to deploy search models
“So that's what we use for, ah, we use to decide which model to push out.”
Assertion Not checkable as stated
Bloomberg Terminal search volume reaches the low hundreds of thousands daily
“I think on a good day we are talking about low 100,000.”