{"id":214692,"date":"2023-12-18T11:53:55","date_gmt":"2023-12-18T06:23:55","guid":{"rendered":"https:\/\/dripp.zone\/news\/mint-explainer-the-mercurial-rise-of-india-focused-llms-crypto-news\/"},"modified":"2023-12-18T11:53:58","modified_gmt":"2023-12-18T06:23:58","slug":"mint-explainer-the-mercurial-rise-of-india-focused-llms-crypto-news","status":"publish","type":"post","link":"https:\/\/dripp.zone\/news\/mint-explainer-the-mercurial-rise-of-india-focused-llms-crypto-news\/","title":{"rendered":"Mint Explainer: The mercurial rise of India-focused LLMs &#8211; Crypto News"},"content":{"rendered":"<p><\/p>\n<div id=\"paywall_11702877613645\">\n<p>      Aggarwal, thus, joins the growing ranks of Indian companies that are building large language models (LLMs) trained on Indian languages. The companies include Bhashini\u2013a unit of the national language translation mission by the ministry of electronics and information technology (Meity); <a rel=\"nofollow noopener\" target=\"_blank\" class=\"autobacklink-topic\" href=\"https:\/\/www.livemint.com\/topic\/tech-mahindra\" data-name=\"Tech Mahindra\">Tech Mahindra<\/a>\u2019s Indus project; AI4Bharat at IIT-Madras; Project Vaani\u2013part of the Bhasha AI project of ARTPARK and the Indian Institute of Science\u2019s pan-India language initiatives; Sarvam AI\u2019s OpenHathi series; and CoRover.ai\u2019s BharatGPT.<\/p>\n<p>      Generative AI, or GenAI, refers to the ability of LLM-powered chatbots such as <a rel=\"nofollow noopener\" target=\"_blank\" class=\"autobacklink-topic\" href=\"https:\/\/www.livemint.com\/topic\/chatgpt\" data-name=\"ChatGPT\">ChatGPT<\/a> to create new content, including audio, code, images, text, simulations, and videos (hence the term, multimodal). GenAI systems fall under the broad category of machine learning, but unlike traditional ML that can analyse data patterns to make predictions, these systems create entirely new content with the help of \u2018prompts\u2019.<\/p>\n<p>      That said, can Ola\u2019s Aggarwal do a Google Gemini or OpenAI\u2019s GPT-4? And why is the parent of electric cars and scooters, Ola Electric, and ride-sharing startup, Ola Cabs, dabbling with foundational models, data centres, and silicon chips that require a lot of investment?<\/p>\n<h2>What\u2019s Krutrim got to do with Ola?<\/h2>\n<p>Aggarwal\u2019s Krutrim announcement comes at a time when the government is set to unveil its AI policy under the India AI programme on 10 January, which will include a policy framework for public-private partnership models on development of AI databases in Indic languages, as well as indigenous compute capacities,\u00a0<a rel=\"nofollow noopener\" target=\"_blank\" style=\"text-decoration:none;\" href=\"https:\/\/www.livemint.com\/ai\/artificial-intelligence\/ola-cofounder-s-ai-venture-keeps-most-details-away-for-the-future-11702651991138.html\">according to Union minister of state for IT, Rajeev Chandrasekhar<\/a>. But the release of Krutrim\u2019s base foundation model also comes at a time when Ola Electric is gearing up to\u00a0<a rel=\"nofollow noopener\" target=\"_blank\" style=\"text-decoration:none;\" href=\"https:\/\/www.livemint.com\/market\/stock-market-news\/ola-electric-to-file-drhp-with-sebi-this-month-to-raise-700-million-via-ipo-report-11702297966173.html\">file for an IPO<\/a>.<\/p>\n<p>      Backed by SoftBank, Ola Electric is targeting a valuation of $7-8 billion by early 2024. While that figure\u2019s much higher than the company\u2019s current estimated worth of\u00a0<a rel=\"nofollow noopener\" target=\"_blank\" style=\"text-decoration:none;\" href=\"https:\/\/www.livemint.com\/companies\/news\/ola-revenue-up-2x-in-fy22-losses-widen-as-costs-rise-11691691997424.html\">about $3.6 billion<\/a>, it\u2019s closer to Ola Electric\u2019s estimated valuation of $7.3 billion as at the end of 2021.\u00a0<\/p>\n<p>      Ola Electric plans to use the funds raised from the IPO for expanding its electric vehicle business and establishing a dedicated lithium-ion cell manufacturing unit.<\/p>\n<p>      Aggarwal has clarified that Krutrim is a \u201cseparate business altogether&#8221;, and will not \u201cbe integrated at a transactional level&#8221;.\u00a0<\/p>\n<p>      \u201cThere are some entities that I own 100%\u2014this is under my company, and not part of Ola or Ola Electric\u2019s corporate structure,&#8221;\u00a0<a rel=\"nofollow noopener\" target=\"_blank\" style=\"text-decoration:none;\" href=\"https:\/\/www.livemint.com\/ai\/artificial-intelligence\/ola-cofounder-s-ai-venture-keeps-most-details-away-for-the-future-11702651991138.html\">he said<\/a>. Aggarwal did say Krutrim had \u201csome investments into (Ola Electric)&#8221;, but did not disclose any details.\u00a0<\/p>\n<p>      Further,\u00a0<a rel=\"nofollow noopener\" target=\"_blank\" style=\"text-decoration:none;\" href=\"https:\/\/www.youtube.com\/watch?v=EP1x_9LMp50\">in a presentation<\/a>, Aggarwal said that all Ola group companies were \u201calready using Krutrim for a lot of their internal workloads, be it customer support, voice and chat, customer sales calls, and for other processes\u2026&#8221;\u00a0<\/p>\n<p>      This clearly implies that Krutrim\u2019s products and services will be cross-sold to enhance the offerings of the group companies.<\/p>\n<h2>How\u2019s GenAI used in vehicles?<\/h2>\n<p>The use of generative AI in the auto sector is not new. Mercedes-Benz, for instance, recently used ChatGPT to power voice assistants in a beta program available to more than 900,000 vehicles.<\/p>\n<p>      Also consider the example of a Formula E electric race car, the GENBETA,\u00a0 an enhanced GEN3 race car. The GEN3 is the fastest, lightest, electric race car with a top speed of more than 322 kmph, and is used by the 11 teams and 22 drivers in the ABB FIA Formula E World Championship.<\/p>\n<p>      Google Cloud provided generative AI to analyse the drivers\u2019 runs. Additionally, experts from McKinsey &#038; Co.\u2019s AI arm, called QuantumBlack, built data and analytics components to create the driver interface that analysed and queried data in real-time using\u00a0<a rel=\"nofollow noopener\" target=\"_blank\" style=\"text-decoration:none;\" href=\"https:\/\/www.mckinsey.com\/capabilities\/quantumblack\/our-insights\/the-state-of-ai-in-2023-generative-ais-breakout-year\">generative AI<\/a>.<\/p>\n<p>      According to Nvidia, generative AI is also enabling new breakthroughs in autonomous vehicle development in research areas including the use of neural radiance field technology to turn recorded sensor data into fully interactive 3D simulations. These digital twin environments, as well as synthetic data generation, can be used to develop, test and validate autonomous vehicles at incredible scale.<\/p>\n<p>      Aggarwal\u2019s AI ambitions, however, appear to go far beyond just the auto sector, given that the Ola group\u2019s businesses extend beyond mobility to financial services offerings including payment systems, insurance agents and cloud kitchens.<\/p>\n<h2>What\u2019s the plan with Krutrium?<\/h2>\n<p>Krutrim\u2019s AI model, according to the company, has been trained on more than 2 trillion tokens (loosely, numerical representation of pieces of words and sub-words that an LLM can understand. For instance, banana is a word, while homework can be split into two words\u2013home and work). While Aggarwal compared Krutrim to GPT4, the latter has been trained on more than 13 trillion tokens.\u00a0<\/p>\n<p>      That said, the strength of Krutrim may lie in its understanding of 20 Indian languages and generating content in 10 Indian languages, including Marathi, Hindi, Telugu, Kannada, and Odiya. Aggarwal said Krutrim has been \u201ctrained on 20 times more Indic tokens than any other model, ensuring a deep understanding of Indian culture, values, and aspirations&#8221;.<\/p>\n<p>      While there\u2019s a waitlist if you register for the base LLM model at\u00a0<a rel=\"nofollow noopener\" target=\"_blank\" style=\"text-decoration:none;\" href=\"https:\/\/olakrutrim.com\/\">OlaKrutrim<\/a>, Aggarwal plans to make the \u201cwhole platform&#8221; available for developers to build application programming interfaces, or APIs, for enterprise applications, in February. Ola also plans to launch Krutrim Pro in the next quarter.<\/p>\n<h2>Can Ola afford a Krutrim?<\/h2>\n<p>That said, building a foundational model from scratch is an expensive affair. OpenAI\u2019s GPT was in the works for more than six years and cost upwards of $100 million and used an estimated 30,000 graphics processing units (GPUs). Aggarwal has not disclosed any details of his investments, or the costs, in Krutrim so far.<\/p>\n<p>      In FY22, Ola derived about 61% of its revenue, or  <span class=\"webrupee\">\u20b9<\/span>1,208.6 crore, from its ride-hailing business in India, while posting a loss of  <span class=\"webrupee\">\u20b9<\/span>101 crore. Financial services comprised a small part of the revenue. The group posted a\u00a0<a rel=\"nofollow noopener\" target=\"_blank\" style=\"text-decoration:none;\" href=\"https:\/\/www.livemint.com\/companies\/news\/ola-revenue-up-2x-in-fy22-losses-widen-as-costs-rise-11691691997424.html\">consolidated operating revenue<\/a> of  <span class=\"webrupee\">\u20b9<\/span>1,970.4 crore in FY22, rising from  <span class=\"webrupee\">\u20b9<\/span>983.2 crore in the year before. Ola\u2019s net losses, though, widened in FY22 to  <span class=\"webrupee\">\u20b9<\/span>1,522.33 crore from  <span class=\"webrupee\">\u20b9<\/span>1,116.6 crore in the previous year.<\/p>\n<p>      That said, since Krutrim is a separate business, Aggarwal may be bootstrapping the venture, given that he has a\u00a0<a rel=\"nofollow noopener\" target=\"_blank\" style=\"text-decoration:none;\" href=\"https:\/\/www.livemint.com\/politics\/news\/startup-founders-sparkle-in-hurun-rich-list-11663785609563.html\">personal net worth<\/a> of a little over $1.4 billion. One, however, will have to wait till Aggarwal discloses more details about his investment plans in designing silicon chips and building the LLM ecosystem.<\/p>\n<h2>How can GenAI work with regional languages?<\/h2>\n<p>The fact remains that even though India is home to more than 400 languages, making it one of the most linguistically diverse countries in the world, most foundation models and LLMs are trained primarily using internet data, which is predominantly English. As per Statista, English was the most popular language for web content, representing nearly 59% of websites as of January this year. Russian ranked second with 5.3% of web content, followed by Spanish with 4.3%.<\/p>\n<p>      While one can only but laud the contribution of India\u2019s Centre for Development of Advanced Computing (C-DAC) in developing the country\u2019s multilingual ecosystem\u00a0<a rel=\"nofollow noopener\" target=\"_blank\" style=\"text-decoration:none;\" href=\"https:\/\/www.livemint.com\/mint-top-newsletter\/techtalk15122023.html\">over the past three decades<\/a>, the fact remains that AI models need to be trained using regional languages to bridge the digital divide in countries like India, which is why efforts such as Krutrim make a lot of sense.<\/p>\n<p>      Krutrim, on its part, says it will tap Bhashini, whose technology comprises automatic speech recognition, optical character recognition, natural language understanding, machine translation, and text-to-speech. The Bhashini platform, for instance, uses optical character recognition (OCR) to extract text from data of printed materials such as brochures to train AI models in 14 languages.\u00a0<\/p>\n<p>      But getting local datasets is a challenge, according to the CEO of Bhashini, Amitabh Nag, who\u00a0<a rel=\"nofollow noopener\" target=\"_blank\" style=\"text-decoration:none;\" href=\"https:\/\/www.livemint.com\/mint-top-newsletter\/techtalk15122023.html\">pointed out<\/a> that many of the 22 official Indian languages do not have digital data, which makes it challenging to build and train an AI model. Bhashini has so far spent $6-7 million to collect data from different sources and employed more than 200 people to collect data (text as well as speech) and feed it into the system, following which the data is curated, annotated, and labelled.<\/p>\n<h2>What other Indic LLMs are in the works?<\/h2>\n<ul>\n<li>The \u2018Nilekani Center at AI4Bharat\u2019 (named after Nandan Nilekani), launched at the Indian Institute of Technology-Madras in July last year, is building open-source language AI for Indian languages, including datasets, models, and applications. The project is supported by EkStep Foundation, Microsoft\u2019s Research Lab, and the India Development Center.<\/li>\n<li>Sarvam AI, a generative AI startup founded by Vivek Raghavan and Pratyush Kumar (both co-founders of AI4Bharat), is developing LLMs specifically for India\u2013<a rel=\"nofollow noopener\" target=\"_blank\" style=\"text-decoration:none;\" href=\"https:\/\/www.sarvam.ai\/blog\/announcing-openhathi-series\">the OpenHathi Series<\/a>. The startup will focus on training AI models to support the diverse set of Indian languages and voice-first interfaces. It will work with Indian enterprises to co-build domain-specific AI models on their data, and also plans to use GenAI atop the India stack (Aadhaar, UPI, Account Aggregator, etc.) \u201cspecifically for public-good applications&#8221;. Sarvam AI is partnering with AI4Bharat, which has \u201ccontributed language resources and benchmarks&#8221;.<\/li>\n<li>Bangalore-based AI and Robotics Technology Park (ARTPARK) and the Indian Institute of Science are partnering with Google India to launch a large language model called Project Vaani. This is part of the Bhasha AI project of ARTPARK and IISc\u2019s pan-India language initiatives, which includes SYSPIN (Synthesizing Speech in Indian languages) and RESPIN (Recognizing Speech in Indian languages). While Google plans to collect speech samples from 773 districts, the initiative is currently focused on 80 districts of 10 states. It is expected to expand over the next couple of years, with over 150,000 hours of curated speech and 100 million sentences of text in Indian scripts.<\/li>\n<li>Cloud-based communications startup Ozontel, too, recently partnered with Swecha Telangana at the Indian Institute of Information Technology-Hyderabad to compile a Telugu stories dataset, aimed at building a Telugu LLM. About 8,000 students from 20 colleges participated to create 40,000 pages of Telugu content.<\/li>\n<li>CoRover has launched its own indigenous LLM called BharatGPT, which is available in more than 12 Indian languages in partnership with Bhashini. CoRover Pvt. Ltd currently offers AI Virtual Assistants (chatbots, voicebots, videobots) to organisations including IRCTC, LIC, the Indian Navy (GRSE), Max Life Insurance, and NPCI. The company is hosted on the Google CloudPlatform (GCP), and Google\u2019s Vertex AI is integrated with CoRover\u2019s conversational AI platform, allowing organisations to utilise Google\u2019s AI services.<\/li>\n<li>And in another effort in the auto sector, the Mahindra Group said in August that it aimed to construct an indigenous LLM specifically designed to converse in a multitude of Indic languages. In the first phase, the Indus Project targets the inclusion of a remarkable 40 Hindi dialects, paving the way for an ever-expanding roster. Tech Mahindra acknowledges it has \u201cdrawn inspiration from \u2018Bhashini\u2019&#8230; to amass datasets on Indic languages&#8221;.<\/li>\n<\/ul>\n<p>   <input type=\"hidden\" id=\"iframecount\" value=\"0\"\/>  <\/div>\n","protected":false},"excerpt":{"rendered":"<p>Aggarwal, thus, joins the growing ranks of Indian companies that are building large language models (LLMs) trained on Indian languages. The companies include Bhashini\u2013a unit of the national language translation mission by the ministry of electronics and information technology (Meity); Tech Mahindra\u2019s Indus project; AI4Bharat at IIT-Madras; Project Vaani\u2013part of the Bhasha AI project of [&hellip;]<\/p>\n","protected":false},"author":2,"featured_media":214693,"comment_status":"open","ping_status":"open","sticky":false,"template":"","format":"standard","meta":{"footnotes":""},"categories":[8],"tags":[15251,15254,15262,15258,15253,11592,10925,188,183,5603,15257,185,186,15260,15259,11391,8392,4664,15265,15261,15249,15250,15263,15176,15177,8714,4665,187,184,15264,15194,10920,5792,15256,15252,15255,189,150,437,182,4667,190],"class_list":["post-214692","post","type-post","status-publish","format-standard","has-post-thumbnail","hentry","category-technology","tag-ai4bharat","tag-artpark","tag-auto-sector","tag-bharatgpt","tag-bhasha-ai","tag-bhashini","tag-bhavish-aggarwal","tag-blockchain-tech","tag-blockchain-technology","tag-chatbot","tag-corover-ai","tag-crypto-technology","tag-cryptocurrency-technology","tag-data-centres","tag-foundational-models","tag-genai","tag-generative-ai","tag-hindi","tag-indian-langugages","tag-indic-languages","tag-indic-llms","tag-indus-projec","tag-kannada","tag-krutrim","tag-krutrim-ai","tag-llms","tag-marathi","tag-metaverse-technology","tag-nft-technology","tag-odiya","tag-ola","tag-ola-electric","tag-openai","tag-openhathi","tag-project-vaani","tag-sarvam-ai","tag-soul-bound-token","tag-tech","tag-tech-mahindra","tag-technology","tag-telugu","tag-token-technology"],"_links":{"self":[{"href":"https:\/\/dripp.zone\/news\/wp-json\/wp\/v2\/posts\/214692","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/dripp.zone\/news\/wp-json\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/dripp.zone\/news\/wp-json\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/dripp.zone\/news\/wp-json\/wp\/v2\/users\/2"}],"replies":[{"embeddable":true,"href":"https:\/\/dripp.zone\/news\/wp-json\/wp\/v2\/comments?post=214692"}],"version-history":[{"count":1,"href":"https:\/\/dripp.zone\/news\/wp-json\/wp\/v2\/posts\/214692\/revisions"}],"predecessor-version":[{"id":214694,"href":"https:\/\/dripp.zone\/news\/wp-json\/wp\/v2\/posts\/214692\/revisions\/214694"}],"wp:featuredmedia":[{"embeddable":true,"href":"https:\/\/dripp.zone\/news\/wp-json\/wp\/v2\/media\/214693"}],"wp:attachment":[{"href":"https:\/\/dripp.zone\/news\/wp-json\/wp\/v2\/media?parent=214692"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/dripp.zone\/news\/wp-json\/wp\/v2\/categories?post=214692"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/dripp.zone\/news\/wp-json\/wp\/v2\/tags?post=214692"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}