Professor, Columbia U., Teach Machine Learning, Cryptocurrency; CTO, Rand Corp; Chief Data Scientist, AIG; Research VP, Xerox; Bellcore/Bell Labs Fellow, Exec D
Fun family story: my dad (@siddharthadalal) taught me Basic programming at 6, C programming at 12. I went to work in tech.
Then I showed him Bitcoin and Ethereum at 67. He now teaches a class on crypto at Columbia.
1/9 🧵 OpenAI's latest model, GPT-4 o1-preview, demonstrates impressive results using Chain of Thought (CoT) reasoning internally. Our recent paper "Beyond the Black Box: A Statistical Model for LLM Reasoning and Inference" explains how and why https://t.co/Zz2fcRRuWx
I just gave a talk at Bangaloru on Artificial General Intelligence with a number of applications in the context of India and Jainism. YouTube link follows. Will appreciate comments. https://t.co/OIgrxKXolU
@agpatriota@vishalmisra@sirbayes There are many ways for constructing priors (e.g., experts-prior experience==training data). So in effect, LLM may be creating prior. By Dirichlet (D) approx any prior on multinomial is a mixture of D, and each D has the behavior of In-Context-Learning. Thus for any prior.
@agpatriota@vishalmisra@sirbayes However, the paper shows that in effect what is generated, as observed in the in-context learning and prompt responses, is consistent with Bayesian Learning with a prior. So it is consistent with implicitly defining prior and updating it.
@agpatriota@vishalmisra@sirbayes As mentioned in the paper, the current generation of LLMs are explicitly using transformers for multinomial probabilities and generating responses. There are other architectures like Mamba that produce similar behavior.
New work by us on Large Language Models - how and why they work, and what is “in-context-learning”
We show that (to quote Dave Blei when he saw our work) “ICL is not magic: it is (consistent with) LM smoothing and Bayesian statistics!” (1/n)
https://t.co/Zz2fcRQX6Z
"I hope the world will get new Gandhis, Mandelas, and Martin Luther Kings," says Siddhartha Dalal, whose gift supports @RutgersU program in Jainism, a faith focused on nonviolence. https://t.co/RZtdtsuWOz @RU_Foundation, @RutgersSASHUM, @RUDiversity, @jainism,
Ransomware attacks are becoming increasingly common and are hard to stop because of the pseudo-anonymity of the Bitcoin network. But Professor Siddhartha Dalal is on the case, with new ways to track down bad actors. @ColumbiaSPS https://t.co/W3ahXlgL5a
I ended my day with this excellent conversation about Leadership in a Time of Global Crises: Overcoming Obstacles with Professor Amy Hungerford in conversation with @SidDalal, @BethFishYoshida, @hedhj16, @Revkin, and Arthur Lerner-Lam.
This event was organized by @Columbia_SPS.
In 2016, ALS tragically snatched away my wife Alka's life as it has for many others. Support my first half marathon in NYC to find a cure to ALS: https://t.co/UV190UdCZA
@fedex Failed to deliver a package promised today - forced me to wait whole day; tracking information was wrong. No specific time for tomorrow's delivery. Don't trust delivery date and time. VERY POOR SERVICE!
Is the media biased against Trump? Researchers @siddalal, Berkay Adlim & @Lesk_M have an answer. Read “How to Measure Relative Bias in Media Coverage” in the October issue of @signmagazine, a joint publication of @AmstatNews and @RoyalStatSoc. https://t.co/xYwTeBBvPy
https://t.co/S3z19fw427
Since the outcomes of quantum computing are probabilistic, it would be great to understand more about uncertainties of the answers.
#cryptocurrency I would appreciate material for a lecture on new regulations and laws on cryptocurrencies in my course on Blockchains/Cryptocurrencies/Analytics at Columbia U.