Technologies: Elasticsearch, OpenSearch, Vespa and practically every DB under the sun. Language agnostic but the bulk of my work has been in Java/Kotlin/Scala, Python, and some Typescript. AWS, NLP, LLM, Transformers.
About me: experienced software engineer passionate about search technologies. Looking to break into proper machine learning (I have experience) and/or build something interesting/meaningful. Long-time backend developer, inexperienced in backend. Been navigating the startup space lately, but it is time to give up and focus on finding something permanent.
This behavior is pretty much the state of the American two party political system. For example, Democrats had countless attempts to make abortion a constitutional right, but if they did, they could no longer count on using that subject for fundraising against the "enemy". Using Democrats as an example, both parties are guilty.
Vector search is not exclusively in the domain of text search. There is always image/video search.
But pre-filtering is important, since you want to reduce the set of items to be matched on and it feels like Elasticsearch/OpenSearch are fairing better in this regard. Mixed scoring derived from both both sparse and dense calculations is also important, which is another strength of ES/OS.
Curious about the lack of Vespa, especially given the thoroughness of the article and its long-time reputation. OpenSearch is also missing, but perhaps it can be considered being lumped in with Elasticsearch due to them both being based on Lucene. The products are starting to diverge, so would be nice to see, especially since it is open-source.
For the performance-based columns, would be also helpful to see which versions were tested. There is so much attention lately for vector databases, that they all are making great strides forward. The Lucene updates are notable.
I spent 3 months in South Africa, primarily in Cape Town. Visted Mabu Vinyl, but sadly did not purchase anything (was backpacking, so traveling light). Owner was lovely.
The documentary is not exaggerating the fact that Rodriguez was immensely popular in South Africa. Heard his music playing on iPods in bars, young 20 somethings that not only listened to him, but so did their parents.
The movie does exaggerate his disappearance. Rodriguez toured Australia, so he knew he had some popularity.
A few years later, I met the owner of Light in the Attic records, the reissue label that repressed Rodriguez's albums. They acquired the rights before the documentary, and even they were amazed at the popularity of his music after the movie. He told basically that Rodriguez paid for his house due to the sales.
That is what I assumed as well, until one day I got hit by a bug involving Extension Mechanisms for DNS (EDNS). Never knew it existed. All of a sudden DNS was failing and could not understand why. Took me a long time to fix the issue.
Since Open Source has been established in the tech ethos for a while now, any deviation has been met with derision. It seems like the community has been more tolerant of these "open" licenses as of late. While must of the hate for projects that do not fit the FOSS standard is mostly unwarranted, hopefully we are not moving quickly in the "open" direction.
The overwhelming majority of users on most social networking sites are lurkers. Those lurkers did not vote. I would assuming the majority of users do not even know about the changes or the blackout.
That was pretty much my experience. Immigrant family and since I went to high school in Europe, I did not have much of a choice for university and ended up at CUNY. Worked evening in supermarkets for the first couple of years, then I was able to get an internship due to my grades, but basically installing computers, not development. There were not many internships available. Eventually went to a top engineering uni for grad school. Despite all of this, I am still considered privileged and white. When you have no circle, no network, no money, when no one in your family went to college, everything is more difficult.
I assumed it had something to do with the color and tried to find patterns. Turns out you need knowledge of the London Metro system. It would be nice if the title reflected that.
My personal definition of Big Data has always been when you gather/store data without having a planned use for it. Do we need this data? Don't know, let's just store it for now.
The article does allude to this definition when it states that "Most data is rarely queried". We have become data hoarders. Technology has made it easy (and relatively cheap) to store data, but the ideas of what to do with this data have not scaled in comparison.
I've been doing the same using Strava Heatmaps, cycling the Hollywood Hills. Lots of very steep dead-end streets. Cannot just wing it since you never know when there might be a street coming up, new streets need to be planned.
Remote: Yes
Willing to relocate: No (for now)
Technologies: Elasticsearch, OpenSearch, Vespa and practically every DB under the sun. Language agnostic but the bulk of my work has been in Java/Kotlin/Scala, Python, and some Typescript. AWS, NLP, LLM, Transformers.
Email: https://url.dev/m/xpQpdwn/
About me: experienced software engineer passionate about search technologies. Looking to break into proper machine learning (I have experience) and/or build something interesting/meaningful. Long-time backend developer, inexperienced in backend. Been navigating the startup space lately, but it is time to give up and focus on finding something permanent.