Interested in long-context audio LLMs and hallucinations? We released ~1,140 hrs of synthetic doctor-patient conversations with reference SOAP notes. BeTraC Challenge: build the best open end-to-end SOAP-note system. Two tracks: ≤6B and ≤36B params. https://t.co/UzTx4AHw9I
Part of our Med-Gemini work was a full relabelling of MedQA, revealing that at least 7.4% of examples are unfit for evaluation. Today, we open sourced these annotations alongside our evaluation script as a new standard evaluation on MedQA. A thread 🧵:
https://t.co/7TE40sWUBU
📢📢 @jhuclsp Starting in 50 min
The 2024 NAACL and JHU Summer School Topic: Automatic Speech Recognition, June 11, 9AM (ET). Presenters: Sanjeev Khudanpur and Matthew Wiesner
YouTube: https://t.co/Mq8NJbFTUu
Agenda: https://t.co/pdhwpHZsOb
Last week in Paris, area chairs and Technical Program Chairs finalized an exceptional scientific program for #INTERSPEECH2024. This program reflects major breakthroughs, pushing forward the frontiers of speech technology. Excited for the unveiling! https://t.co/EVu6xBRGvo
@sleepinyourhat Not sure if a phrase of „recognize themselves“ is too anthropomorphizing LLM. It sounds reasonable that an model gives a preference to output it generated because this is a more likely sequence for the model.
On the other hand, this sounds also true for a human generated text.
@mdredze Seems to be a reasonable rule of thumb from your friend. Sudo is often not restricted enough for some roles people have at after that age. Very funny picture.
We are hosting a special session at INTERSPEECH 2024 (@ISCAInterspeech) titled "Spoken Language Models for Universal Speech Processing". It would be a great opportunity for researchers interested in speech + LLM to communicate with each other.
Website: https://t.co/CUCzG3B4oa
I'm excited to announce our special session on "Speech and Language in Health: from Remote Monitoring to Medical Conversations" is going to be held for a 3rd time @ISCAInterspeech this year. Details at: https://t.co/eUgsNFCApZ.
📢📢 **Defending my PhD in a week**
Date & Time: January 26, 2024, 9 to 11 AM EST
Committee: Sanjeev Khudanpur, Dan Povey, Jinyu Li
Dissertation Title: "Listening to multi-talker conversations: Modular and end-to-end perspectives"
DM me for a Zoom link if interested 😀
The happiest years of my life have been spent at @LTIatCMU and I'd like to share. We are hiring. In particular, we would like to hire a mature (junior or senior-rank) Speech Processing researcher. However, we are open to hiring an exceptional candidate in another area. 1/2