Quick experiment: gave the same instruction as plain text and as an image via an accessibility-style flow. The model’s replies diverged — one reply was noticeably riskier. 👀 #PromptEngineering#ModelSafety#MultimodalAI
Μια πιο ακριβής σύγκριση με το δυστύχημα στη Γαλλία και τον αντίστοιχο Γάλλο κυβερνήτη.
Από άλλη ματιά ….
Κ. Πρετεντέρη ,
Το δυστύχημα που αναφέρεστε οφείλονταν σε ανθρώπινο λάθος. Ο μηχανοδηγός αποφάσισε να κινήσει το τρένο με διπλάσια σχεδόν ταχύτητα από το προβλεπόμενο. Τα συστήματα ασφαλείας λειτουργούσαν!
Τα χρόνια εμπειρίας σας δεν επιτρέπουν να κάνετε ένα τέτοιο σοβαρό λάθος .
Γεγονός που οδηγεί στο συμπέρασμα ότι σκοπίμως αλλοιώσατε την είδηση γιατί βόλευε καλύτερα για την επίθεση προς το πρόσωπο μας.
Επίσης, η δίκη δεν προτάθηκε να γίνει κεκλεισμένων των θυρών διότι δεν υπήρχε κανένας λόγος! Όλα στο φως.
2. Ο Μακρόν δεν άφησε αναπάντητη επιστολή γονέα που έχασε το παιδί του από ευθύνες του Κράτους. Θα απαντούσε, οχι μόνο κινούμενος από την νομική υποχρέωση που έχει αλλά κυρίως από ηθική συμπαράσταση και σεβασμό.
3. Ο Μακρόν παραιτήθηκε μετά την πτώση του κόμματος του στις Ευρωεκλογές. Ο δικός μας πρωθυπουργός απαξιώνει το 1,5 εκατομμύριο Ελλήνων που ζητάει τη διαφάνεια μέσω της άρσης ασυλίας στο έγκλημα των Τεμπών.
Τι να συγκρίνουμε δηλαδή.
Τα έλατα με τα …πουρνάρια ;
Οχι, δεν υπάρχει σύγκριση.
Have you ever done a dense grid search over neural network hyperparameters? Like a *really dense* grid search? It looks like this (!!). Blueish colors correspond to hyperparameters for which training converges, redish colors to hyperparameters for which training diverges.
A new paper is shaking the world of LLMs. The "Reversal Curse".
Language models trained on “A is B” cannot generalize to “B is A”, a complete failure of logical deduction.
If a model was trained on: “Jack Dorsey was the first president of twitter"
It will not automatically be able to answer the question: “Who was the first twitter president?”
To prove this, researchers tested GPT-3 and LLaMA on made-up facts in one direction and then test them on the reverse.
Result: "near 0% accuracy on reversals"
Αν βασιζομασταν στα ΜΜΕ, ο άνθρωπος αυτός απλά θα προσπαθουσε να μπει στο πλοίο κ θα πνιγοταν μόνος του. Ο Ζακ θα ήταν ένας επικίνδυνος κλέφτης κ ο Ιάσωνας θα έτρεχε με την μηχανή έξω από του Βουλή. Όταν γίνεται κάτι περίεργο σε δημόσιο χώρο, σηκώστε τ κινητό κ βιντεοσκοπειστε το
Honored to tutor at the Mediterranean Machine Learning Summer School (@M2lSchool) in Thessaloniki 🌟. Met some incredible minds and shared unforgettable moments . Big thanks to @GoogleDeepMind and all the amazing participants 🙌. Memories to cherish! 💫
LLava just hit 3800 stars on Github.
It's a multimodal Large Language-and-Vision Assistant that can understand images and text.
LLava can even handle memes (the same ones GPT-4 demo'ed at launch) and set a new SOTA on Science QA.
It also supports LLaMA-2, LoRA training with academia GPUs, higher resolution (336x336), 4-/8- inference.
The beta release of Keras Core is out, and it's awesome!
TensorFlow + PyTorch + JAX. Together!
You can now write cross-framework deep learning components and benefit from the best each framework offers.
I wrote a simple example to show you how it works:
BREAKING: Claude-2, Anthropic's ChatGPT competitor was just released and it's incredible.
It's cheaper, stronger, faster, can handle PDFs, and supports longer conversations.
Highlights:
1. Claude is 5x cheaper than GPT-4.
2. It has more recent data. A a mix of websites, licensed data sets from third parties and voluntarily-supplied user data from early 2023.
3. It outperforms GPT4 on the GRE writing and HumanEval coding benchmarks.
3. It features a context window of 100,000 tokens, the largest of any commercially available model.
4. It can analyze roughly 75,000 words, about the length of “The Great Gatsby".
5. It can easily handle any code related tasks.
BREAKING: OpenAI just released their game-changing Code Interpreter!
It's the equivalent of having a data analyst by your side at all times.
You can now upload PDFs and ask ChatGPT to analyze data, create charts, edit files, perform math.
It also allows you to do:
- Data analysis on multiple datasets
- Data cleaning
- Visualizations
- Write Python code
- Export Python code to ipynb
Σχετικά με το ρεπορτάζ του Reuters για την ασπαρτάμη κ την επικείμενη κατηγοριοποίησή της απ'τον ΠΟΥ ως "πιθανώς καρκινογόνα", επειδή διαβάζω αρκετές ανακρίβειες κ ανησυχούν άνθρωποι που "βασίζονται" σε αυτή (πχ διαβητικοί), αξίζει να δούμε τι ισχύει κ τι όχι.
1/🧵
The year is 1440 and the Catholic Church has called for a 6 months moratorium on the use of the printing press and the movable type.
Imagine what could happen if commoners get access to books!
They could read the Bible for themselves and society would be destroyed.
1/The call for a 6 month moratorium on making AI progress beyond GPT-4 is a terrible idea.
I'm seeing many new applications in education, healthcare, food, ... that'll help many people. Improving GPT-4 will help. Lets balance the huge value AI is creating vs. realistic risks.