Hacker News
new
past
comments
ask
show
jobs
points
by
Colegno
6 hours ago
|
comments
by
quixoticaxolotl
6 hours ago
|
next
[-]
They are already solving the problem with search engines, they're just using an LLM as a first pass to create better embeddings to run a similarity match on first. The difference in latency is likely made up for in accuracy.
reply
by
fastball
6 hours ago
|
prev
|
next
[-]
A 2s LLM call is pretty slow.
reply
by
gadflyinyoureye
6 hours ago
|
parent
|
[-]
Try using Digital Ocean. Minutes spent on inference.
reply
by
nullsanity
6 hours ago
|
prev
|
[-]
[dead]
reply