◈Bluesky100
YHacker News100
XX / Twitter100
project rollup
jev
300
mentions
across 2 keywords
across 2 keywords
net sentiment
+14%
95 positive
151 neutral
54 negative
volume & sentiment over time
Sep 18–26, 2026 · Daily
sources
most discussed
decision / models
active days
9
mentions
Showing 1–25 of 100 mentions
-
Ollaya – Ollama for open-source, Jev-style decision modelsI've been following jevbench twice a day for the past week and that's been a lot of fun. Latest update: Rank System Score Public / sealed accuracy Evidence 1 decider-4b v2 64.13 83.5% / 34.7% Evaluator-run, offline 2 Jev 1.13 63.29 86.6% / 36.7% Evaluator-run API 3 JevK5 v0.2 62.04 85.3% / 33.1% Evaluator-run 4 Cygnet … I've been following jevbench twice a day for the past week and that's been a lot of fun. Latest update: Rank System Score Public / sealed accuracy Evidence 1 decider-4b v2 64.13 83.5% / 34.7% Evaluator-run, offline 2 Jev 1.13 63.29 86.6% / 36.7% Evaluator-run API 3 JevK5 v0.2 62.04 85.3% / 33.1% Evaluator-run 4 Cygnet 12B 61.76 87.9% / 33.8% Evaluator-run, offline 5 Hopper 59.43 82.3% / 34.1% Evaluator-run 28 Kev 4B 36.14 66.2% / 22.4% Evaluator-run 41 Laya 421M 30.25 58.4% / 30.8% Evaluator-runview source ↗
-
Jev BenchmarksThere are comming more and more Jev alternatives. Is there anywhere a benchmarklist of these jev competitors? There are comming more and more Jev alternatives. Is there anywhere a benchmarklist of these jev competitors?view source ↗
-
Ollaya – Ollama for open-source, Jev-style decision modelsIsn't the point of Jev that it generalises better? It's a fast classifier you can use out-the-box, ~1.5bn tokens is about $40 (I've been hammering it) It just works ... a whole bunch of low-level/low-importance workflow stuff that was getting farmed out to small/fast LLM models now has a competitive alternative ... and… Isn't the point of Jev that it generalises better? It's a fast classifier you can use out-the-box, ~1.5bn tokens is about $40 (I've been hammering it) It just works ... a whole bunch of low-level/low-importance workflow stuff that was getting farmed out to small/fast LLM models now has a competitive alternative ... and bits that hadn't even been considered to go into some external descision/classifier service can be tested/deployed at ~$0.00003/req I don't get this wall of negativity on it, it's genuinely innovative/useful tech ... would expect HN to be more positive, regardless of whether it's the absolute best executionview source ↗
-
Dutch governments builds alternative for Microsoft based on NixOSThis, too, gives me vibes of cryptobros! Proponent: EmergingTech is gonna be amazing, it'll at revolutionize everything! Bystander: Alright. But I'll hold off until I see actual evidence of that happening. Truly revolutionizing tech usually shows clear signs of being revolutionary, and doesn't need evangelists. Propone… This, too, gives me vibes of cryptobros! Proponent: EmergingTech is gonna be amazing, it'll at revolutionize everything! Bystander: Alright. But I'll hold off until I see actual evidence of that happening. Truly revolutionizing tech usually shows clear signs of being revolutionary, and doesn't need evangelists. Proponent: You're gonna be left behind! Act now, or forever miss out! Here's analogies with things we know in hindsight were revolutionary. Bystander: Those analogies are bad for these specific reasons. Proponent: You will be proven wrong! It's so dumb. At least you AI proponents don't post pictures of yourself with glowing laser eyes when you evangelize, so I guess that's an improvement over the cryptobros. PS: I emphasize that I'm not claiming that the technology won't have huge effects on society. If that's your take on what I write, you have to read again.view source ↗
-
Ollaya – Ollama for open-source, Jev-style decision models<<<"i was curious to see if i could train a competitive Jev-like model completely autonomously with a swarm of agents using our internal system." Bro is writing off the H200 lol On a sidenote I really can't stand the term "swarm" and definately plays into AI doomerism. <<<"i was curious to see if i could train a competitive Jev-like model completely autonomously with a swarm of agents using our internal system." Bro is writing off the H200 lol On a sidenote I really can't stand the term "swarm" and definately plays into AI doomerism.view source ↗
-
Ollaya – Ollama for open-source, Jev-style decision modelsJev does it more efficiently because it doesn't use an LLM https://typesafe.ai/blog/introducing-system-one-models-and-j... Jev does it more efficiently because it doesn't use an LLM https://typesafe.ai/blog/introducing-system-one-models-and-j...view source ↗
-
Ollaya – Ollama for open-source, Jev-style decision modelstext classification is equivalente to decision. This is exactly the same thing Jev does. text classification is equivalente to decision. This is exactly the same thing Jev does.view source ↗
-
Ollaya – Ollama for open-source, Jev-style decision modelsFair, and I'd be happy if they did. Ollaya uses the same API as Jev, so your code isn't tied to it either way Fair, and I'd be happy if they did. Ollaya uses the same API as Jev, so your code isn't tied to it either wayview source ↗
-
Ollaya – Ollama for open-source, Jev-style decision modelsDepends on the model. The small ones I support today are well below Jev on harder queries, but fine for simple, well-defined questions. The open models that get close to Jev are bigger, and I'm adding support for those next. Depends on the model. The small ones I support today are well below Jev on harder queries, but fine for simple, well-defined questions. The open models that get close to Jev are bigger, and I'm adding support for those next.view source ↗
-
Ollaya – Ollama for open-source, Jev-style decision modelsThe link rgbrgb posted is a good overview. The best open ones are close to Jev now, but they're big models. And I agree, if you have an eval set for a fixed task, a trained classifier is the better choice. The link rgbrgb posted is a good overview. The best open ones are close to Jev now, but they're big models. And I agree, if you have an eval set for a fixed task, a trained classifier is the better choice.view source ↗
-
Ollaya – Ollama for open-source, Jev-style decision modelsYes. JEV generalizes better because they probably have an enormous corpus and trained on it for a long time. Laya's out of the box model is much weaker. However, in the age of LLM's it's incredibly easy and cheap to generate large datasets to fine tune laya for your task, and the training loop is pretty quick and cheap… Yes. JEV generalizes better because they probably have an enormous corpus and trained on it for a long time. Laya's out of the box model is much weaker. However, in the age of LLM's it's incredibly easy and cheap to generate large datasets to fine tune laya for your task, and the training loop is pretty quick and cheap too. It's so easy that I question why I would ever pay for JEV when eventually I'll have done enough random things that I will also have a large corpus and likely a general model as well.view source ↗
-
Ollaya – Ollama for open-source, Jev-style decision modelsDeveloper here. You're right, Laya is a lot weaker than Jev, especially on harder queries. It's a small model, so it's fast, but that's the trade-off. The open models that get close to Jev are much bigger, and running those is what I'm working on next. Developer here. You're right, Laya is a lot weaker than Jev, especially on harder queries. It's a small model, so it's fast, but that's the trade-off. The open models that get close to Jev are much bigger, and running those is what I'm working on next.view source ↗
-
Ollaya – Ollama for open-source, Jev-style decision modelsHas anyone actually seen better or the same results with Laya compared to Jev? From my experience, Laya performs significantly worse. It's less confident and often makes wrong decisions with more complex queries. Has anyone actually seen better or the same results with Laya compared to Jev? From my experience, Laya performs significantly worse. It's less confident and often makes wrong decisions with more complex queries.view source ↗
-
Ollaya – Ollama for open-source, Jev-style decision modelsSounds good on latency but how is its actual decision quality vs. Jev? Sounds good on latency but how is its actual decision quality vs. Jev?view source ↗
-
Ollaya – Ollama for open-source, Jev-style decision modelsI am fairly confident if Jev-style decision models are seen as prominent (which, they seem to be), Ollama will support them. Surprised the team hasn't implemented this already. I am fairly confident if Jev-style decision models are seen as prominent (which, they seem to be), Ollama will support them. Surprised the team hasn't implemented this already.view source ↗
-
Drex: Open jev-like claims win on decision indexview source ↗
-
Ollaya – Ollama for open-source, Jev-style decision modelsthere's this thing with a bunch of similar models https://huggingface.co/spaces/multimodalart/jev-decision-ind... top open one is trained by perplexity cto for $3k, kinda cool https://x.com/denisyarats/status/2102252088067850507 there's this thing with a bunch of similar models https://huggingface.co/spaces/multimodalart/jev-decision-ind... top open one is trained by perplexity cto for $3k, kinda cool https://x.com/denisyarats/status/2102252088067850507view source ↗
-
Jevmem – automatic project memory for Claude Code, built on JevThe people that run local jev like models do sth similar. Train a seperate nn adapter on top of qwen 4b so it has balanced probabilities whatever that means and apparently it works quite well. Side note: conspiracy theorists say that jev is a qwen model fine tuned but who know if true The people that run local jev like models do sth similar. Train a seperate nn adapter on top of qwen 4b so it has balanced probabilities whatever that means and apparently it works quite well. Side note: conspiracy theorists say that jev is a qwen model fine tuned but who know if trueview source ↗
-
Ollaya – Ollama for open-source, Jev-style decision modelsAre there many models that are comparable to Jev for generic decision making? Smarter move if you have an eval set is to just train a classifier and call it a day. Are there many models that are comparable to Jev for generic decision making? Smarter move if you have an eval set is to just train a classifier and call it a day.view source ↗
-
Ollaya – Ollama for open-source, Jev-style decision modelsview source ↗
-
Jev for Input Validationview source ↗
-
Jevmem – automatic project memory for Claude Code, built on JevI really don't undertstand the jev-hype. Just start a cheap chinese agent in a context where it has only one tool, and JSON-schema constrain that toolcall to the response shape you want. Prompt the agent to make only one tool call and not speak. Done. I really don't undertstand the jev-hype. Just start a cheap chinese agent in a context where it has only one tool, and JSON-schema constrain that toolcall to the response shape you want. Prompt the agent to make only one tool call and not speak. Done.view source ↗
-
Jev in Weather ForecastingWeather apps have been notoriously pretty bad when it comes to predicting the vector rain will take on a radar. Dark Sky (bought by Apple) tried to do this with machine learning and was objectively the best for decades. However, given Jev's fast speed + cheap costs, hooked in Jev as the decision maker for forecasting r… Weather apps have been notoriously pretty bad when it comes to predicting the vector rain will take on a radar. Dark Sky (bought by Apple) tried to do this with machine learning and was objectively the best for decades. However, given Jev's fast speed + cheap costs, hooked in Jev as the decision maker for forecasting rain. The resulting lift is meaningful. Made an app on the App Store + published a bit of the work. If there's anyone smart out there, feel free to build on top of it and Jev is just an awesome model to play around with...definitely unlocks use cases that even GPT Luna was too expensive / not rigid enough for.view source ↗
-
Jev in Weather Forecastingview source ↗
-
Jev vs. Kev: open-source Jev alternative tested side by sideview source ↗