The next big opportunity in software will be C and C++ coders who can write world class firmware. With Robotics, Automation, EVs, Space Tech booming in India, we will need finest coders for bare metal coding or application development. Right now, there is a big void in this space.
A couple reflections on the quantum computing breakthrough we just announced...
Most of us grew up learning there are three main types of matter that matter: solid, liquid, and gas. Today, that changed.
After a nearly 20 year pursuit, we’ve created an entirely new state of matter, unlocked by a new class of materials, topoconductors, that enable a fundamental leap in computing.
It powers Majorana 1, the first quantum processing unit built on a topological core.
We believe this breakthrough will allow us to create a truly meaningful quantum computer not in decades, as some have predicted, but in years.
The qubits created with topoconductors are faster, more reliable, and smaller.
They are 1/100th of a millimeter, meaning we now have a clear path to a million-qubit processor.
Imagine a chip that can fit in the palm of your hand yet is capable of solving problems that even all the computers on Earth today combined could not!
Sometimes researchers have to work on things for decades to make progress possible.
It takes patience and persistence to have big impact in the world.
And I am glad we get the opportunity to do just that at Microsoft.
This is our focus: When productivity rises, economies grow faster, benefiting every sector and every corner of the globe.
It’s not about hyping tech; it’s about building technology that truly serves the world.
Here we go.
IndiaAI Mission invites proposals from startups, researchers and entrepreneurs to build state-of-the-art foundation models for the country.
The models can be Large Language Models (LLMs) or Small Language Models (SLM), across multimodal text, voice or video.
They must be indigenous AI models that align with global benchmarks, while addressing unique Indian challenges such as cultural, linguistic & societal context.
Everyone is welcome to apply, multiple proposals will be selected with the goal of having SOTA ready India models by the end of the year.
Proposals need to be sent to tenders[at]indiaai[dot]gov[dot]in.
Time to put India on the map! 🚀🇮🇳
@OfficialINDIAai@GoI_MeitY
https://t.co/gRnmM2hWHt
Writing this as an Indian who works on AI in leadership role for one the largest companies in the world (though strictly my personal opinion, but based on verifiable data).
You heard it first here:
—————————-
First some more shocks:
You heard DeepSeek.
Wait till you hear about Qwen (Alibaba), MiniMax, Kimi, DuoBao (ByteDance) all from China.
Within China, DeepSeek is not unique and their competition is close behind (not far behind).
IMHO, China has 10 labs comparable to OpenAI/Anthropic and another 50 tier 2 labs.
The world will discover them in coming weeks in awe and shock.
AI is not hard (I am not high)
————————————
Ignore Sam Altman.
Many teams that built foundation models are below 50 persons (e.g. Mixtral).
In AI, LLM science part is actually quite easy.
All these models are “Transformer Decoder only models”, an architecture that was invented in late 2017.
There are improvements since then (flash attention, ROPE, MOE, PPO/DPO/GRPO), but they are relatively minor, open source and easy to implement.
Since building foundation models is easy and Nvidia is there to help you (if not directly, then by sharing their software like “Megatron” that is assembly line to build AI models) there are so many foundation models built by Chinese labs as well as global labs.
It is machines that learn by themselves…if you give them data & compute. This is unlike writing operating system or database software. Also, everyone trains on same data: internet archives, books, github code for the first stage called “pre-training”.
What is part is hard then?
———————————-
It is the parallel & distributed computing to run AI training jobs across thousands of GPUs that is hard. DeepSeek did lot of innovation here to save on “flops” and network calls. They used an innovative architecture called Mixture of Experts and a new approach called GRPO. with verifiable rewards both of which are in open domain through 2024.
Also, there is lot of data curation needed particularly for “post training”
to teach model on proper style of answering (SFT/DPO) or to teach them learn to reason (GRPO with verifiable reward). STF/DPO is where “stealing” from existing models to save cost of manual labor may happen.
LLM building is nothing that Indian engineers living in India cannot pull off. Don’t worry about Indians who have left. There are plenty in the country as of today.
Then why India does not have foundation models?
———————
It is for the same reason India does not have Google or Facebook of its own.
You need to able to walk before you can run.
There is no protected market to practice your craft in early days. You will get replaced by American service providers as they are cheaper and better every single time. That is not the case with Chinese player. They have a protected market and leadership who treats this skillset as existential due to geopolitics.
So, even if Chinese models are not good in early days they will continue to get funding from their conglomerates as well as provincial governments. Darwinian competition ensures best rise to the top.
Recall DeepSeek took 2 years to get here without much revenue. They were funded by their parent. Also, most of their engineers are not PHDs.
There is nothing that engineers who built Ola/Swiggy/Flipkart cannot build. Remember these services are second to none when you compare them to their Bay Area counterparts. Also , don’t trivialize those services; there is brilliant engineering to make them work at the price points at which they work.
Indian DARPA with 3B USD in funding over 3 years
———————-
What we need is a mentality that treats this skillset as existential. We need a national fund that will fund such teams and the only expected output will be benchmark performance with benchmarks becoming harder every 6 months . No revenue needed to survive for first 3 years.
That money will be loose change for GOI and world’s richest men living in India.
@protosphinx@balajis@vikramchandra@naval
A new chapter begins today.
In view of the various challenges and opportunities facing us, including recent major developments in AI, it has been decided that it is best that I should focus full time on R&D initiatives, along with pursuing my personal rural development mission.
I will step down as CEO of Zoho Corp and take a new role as Chief Scientist, responsible for deep R&D initiatives. Our co-founder Shailesh Kumar Davey will serve as our new group CEO. Our co-founder Tony Thomas will lead Zoho US. Rajesh Ganesan will lead our ManageEngine division and Mani Vembu will lead the https://t.co/ydhRPSqDRN division.
The future of our company entirely depends on how well we navigate the R&D challenge and I am looking forward to my new assignment with energy and vigor. I am also very happy to get back to hands on technical work. 🙏
🚨🚨 Execs from SoftBank, OpenAI and Oracle expected to say they plan to commit $100 billion initially and pour up to $500B into Stargate over the next four years, per multiple sources. Other details of new partnership not immediately available. Trump scheduled to speak at 4 pm.
SCOOP: President Trump is set to announce billions of dollars in private sector investment to build artificial intelligence infrastructure in the United States, @CBSNews has learned.
OpenAI, Softbank and Oracle are planning a joint venture called Stargate, according to multiple people familiar with the deal.
https://t.co/tgO7VWHddU
disappointing outcome, but just want to thank our community & Figma’s community for all the support and excitement re: the possibilities. so much respect for @zoink and the whole @figma team - they have such a bright future, and I know we’ll look for ways to partner to bring some of our ideas to life.
aside from this process, it’s been a wild 2023 w/ a ton of innovation across our products. looking ahead, the pipeline is strong for the next gen of digital experiences that is more engaging than we can imagine. grateful to be on this journey with the world’s creators.
https://t.co/zgxHuagkQg
At a Dead Sea resort that has become their temporary home, surviving members of Kibbutz Be’eri are grappling with their losses from the Oct. 7 attacks and trying to find a way forward https://t.co/fFYYXnzxVg
🌐 The Transformative Role of Large Language Models (LLMs) in Modern Service Marketplaces & Professional Service/Consulting Firms: Exciting Times Ahead! #GenerativeAI#ProfessionalServices#Consulting https://t.co/lKowc67TFK
We are proud to be the first company in the web design industry and @webflow to have a full components UI library powered by AI!
@Windbase_io library will come with abstract feature images based on @OpenAI's DALL·E 2.
More details coming soon! ⚡️🙏🏼