upvote
It seems like jev's major advantage over existing classifiers is that I dont have train it.

If I have to gather and tag data to fine-tune Jev, I can probably just train an "old school" classifier model and make it even cheaper, faster, and just as accurate.

reply
Also one of the more interesting features of Jev is the confidence rating that hardly any Jev-cc talks about.
reply
It seems interesting to me only in the sense of being useless. From the horse’s mouth:

> Confidence is derived from the probabilities

https://docs.typesafe.ai/confidence

(Why is it much easier to find AI-slop websites quoting this than it is to find the actual documentation?)

My inner Bayesian would like for Jev to provide something resembling “evidence”, although I admit that one might ask Jev questions that are somewhat awkward to treat as typical Bayesian questions. If I ask “will this PR be merged”, it’s kind of strange to contemplate the probability of a PR conditioned in that PR being merged in the future. But I bet there is a way to formalize a prior-free classifier in a way that makes Bayesians and non-Bayesians happy, possibly involving actual learned probabilities and confidence levels. If you read the literature on scoring rules, you will find that classifier scores do somewhat naturally decompose into a few interpretable terms.

reply
I think better than priors would be a closed loop where you tell it what the right answer was (or some signal) and they monitor and fine-tune for you
reply
We were able to completely automate 20,100 token prompts with At0m[https://at0m.pienomial.com/].

We believe entire compliance workflows (even multilingual) could be automated.

Would you like to get a demo ?

reply
You’re coming on a bit strongly - a couple of comments with a link is sufficient. Before trying to sell, try to genuinely further the conversation, provide some useful knowledge in return for the reader’s attention.
reply
Point taken. Let me explain if you allow me.

It is possible with deterministic decision models, such as At0m, to gauge the probabilities at every decision. This behavior in addition to hard coded logic, it is possible to completely replicate a prompt's logic.

Using Fable 5.1, it is a matter of minutes.

I believe that most of the compliance check documents will be a solved problem, 3-6 months in future.

None of the LLMs can do it.

Hence I asked to the comment poster if he would want to demo, so that I can show it to him, how to do it step by step. By bad, if it came out too strongly.

reply