“So do not dare propose that ‘Israel Never Truly Wanted Peace with Palestine’. If anything, we wanted it too much. So now, it is your turn to answer the question: When did you ever truly want peace with the one Jewish state?”
(my 10 min speech with English captions from the Oxford Union Debate):
@rish_gpt@claudeai@OpenAI@dkroshow Agree. Benchmarks still don't cover this well.
if you primarily talking about code then:
https://t.co/PPMI3PKk6l is interesting attempt
there's also new papers called FixedBench and FeatBench that are relevant
for just "chat" i believe this is solvable with system prompt/etc.
@TalMorgenstern לא יודע אם מספיק כדי להשפיע על התמונה הגדולה של תעסוקה, אבל אם תהיה עלייה משמעותית בפנאי, תהיה עליה בתעשיות תומכות פנאי מעבר לתיירות ואטרקציות (כושר, מכשור רפואי, פיתוח תרופות, הארכת חיים) שבהן הבעיות עדיין קשות מא��ד, גם עם AI
@TalMorgenstern זה דווקא מחזק את פייסבוק/אינסטגרם (לעומת טיקטוק) כי הם מתבססים על הסושיאל גרף. או שאתה טוען שגם אנשים שהם לא משפיענים, בתוכן ש״מטרגט״ חברים שלהם, יזייפו/ישתמשו בAI?
לא משתמש ��בל לדעתי סנאפ מפרידים (בבירור יותר) בין ברודקסט (סרטים/טלויזיה) לחברים/סושיאל
@therealnirs אליינמנט כמו שתיארת בהתחלה זה כל שלב הפיין טיון שהוא הרבה מעבר לסייפטי (לענות באורך מסוים כשמבקשים ממך, איך לעשות סיכום, לכתוב אימייל, לכתוב קוד, לדבר בצורת צ׳אט וכו׳ וכו׳). סייפטי, שזה בעיקר מה שציינת, הוא רק חלק מזה
@MattEnthoven As an example of the change in ai -
Today you can couple LLM with the vision model to reason through driving decisions. That was unthinkable in 2018
Re/ how much/big - I don’t know and that’s obv the main q
How better cruise/Waymo need to get to drop teleportations?
Waymo's research notes that understanding the "why" and "what" of situations in training data is crucial to building such a model, which is only able to identify correlations in the data. They also note that understanding the "why" of their results makes it easier to improve said system, which is difficult to do in a "black box" AI model. This is crucial for such a safety critical system.