Localized reading
Jev VS Fable 5.1 VS GPT-6 Astra: 체스 typesafe의 새로운 Jev V13은 LLM이 아닙니다. 채팅도, 설명도, 코드 작성도 하지 않습니다. 단지 결정만 내릴 뿐입니다. 그래서 우리는 프론티어 LLM을 상대로 공세를 펼쳤습니다. 테...
Original post / EN
Jev VS Fable 5.1 VS GPT-6 Astra: chess typesafe's new Jev V13 isn't an LLM. it doesn't chat, doesn't explain, doesn't write code — it only makes decisions. so we made it play blitz against frontier LLMs. the test: 5+0 blitz. every move is one API call. Jev V13 vs Fable 5.1: • fable outplayed it. by move 29 it was +16 in material and even promoted a second queen • but it kept burning 6-15 seconds per move on analysis. jev answered in ~2.6s • so fable didn't have enough time and lost Jev V13 vs Astra: • astra didn't bother winning on material. it mated jev in 18 moves. Qe1#, with 2:27 to spare
- views
- 52.6만
- likes
- 2.2천
- saves
- 749
- reposts
- 138
Quoted post
Diogo Almeida @CompleteSkeptic
After co-inventing ChatGPT, I kept asking myself: why have superhuman chat models not led to AGI? I’ve spent the last 2 years in stealth building a new way to train models (RLCD), and a new type of frontier AI model that we are releasing today: Jev • 20-200x faster • 40-400x https://t.co/JSybNG2BKJ
원문 보기 ↗출처
This independent archive preserves a public post with attribution. Text, media, account details and trademarks belong to their respective owners.
