A few months ago, a post on Reddit became very popular in which a 19-year-old student from Bihar claimed that he was going to build a multimodal AI model with 5.82 billion parameters all by himself. Very few people believed him; some people thought it was a scam and an anonymous user even bought a typo domain in order to set up a smear site called "The Dossier", stating that the boy was a complete fraud. Nevertheless, while it took them weeks to write up the negative articles about him, the boy stayed quiet and reached the age of 20, carrying on with his work.
Hello, I'm the boy in question. My name is Abhinav Anand and I'm from Bihar.
Introducing Arcle V1:
Arcle V1 is a unified omni foundation model consisting of 5.84 billion parameters and is able to handle text, images, documents, speech, and audio while at the same time generating text, original images of dimensions 512×512, and natural speech at 24kHz. The model has an architectural context window of 2,097,152 tokens (2 million tokens), supports functioning in 18 or more languages, showing a special strength in Hindi and having extensive knowledge of India, and achieves \*\*80.0% on ARC-Easy, 77.5% on GSM8K, 74.2% on MATH-500, 67.0% on HellaSwag, and 94.6% document content-word recall\*\*.
It was made, built, trained and tested in Bihar.
Arcle V1 is NOT a Router, NOT a Wrapper, NOT a Pipeline, NOT an API-Based App
Currently, most systems promote themselves as multimodal even though in reality they are just four or five individual models connected together through an API. Arcle V1 is instead a single neural network module; whatever the input may be—whether it is text, vision, documents, speech or audio—all of these are mapped into one common 2,560-dimensional semantic latent space and pass through one set of weights, one model file, and go through a single forward pass.
Arcle V2 Upcoming Features
Arcle V1 is merely the starting point; for Arcle V2 we are planning to engineer:
The omni graph includes full integration of the capability to convert text into video and images into video.
\* \*\*Multi-Voice Expressive Speech Output:\*\* This is a type of speech synthesis which incorporates natural emotions by means of using multiple voices.
There are higher benchmark scores, with considerable improvements in the areas that involve complex mathematical and logical reasoning.
It has a strong capability in the field of cybersecurity, with built-in checks for vulnerabilities, code auditing, and protection against insecure inference.
It is extremely efficient in that it provides on-device inference with a high throughput which has been optimised for local consumer hardware.
For Help: Donations and Why We Need Them
There is considerable difficulty in creating an independent and open AI without the help of corporate venture capitalists; I have already invested all that I have in Arcle V1, and in order to bring about Arcle V2 we need the support of the community.
-
\*\*Calculate / Fund Donations:\*\* The money raised by the community is sent directly to the GPU compute clusters so that the Arcle V2 training run can take place.
-
\*\*Donations of data (codebases, technical PDFs, books):\*\* Some AI companies scan rare historical books and proprietary knowledge and then regard human heritage as their own, offering it out only via closed subscription services. Our objective is to preserve this knowledge and make it accessible to all. We would be pleased if you donated your technical PDFs, codebases, and scanned literature. \*\*Our promise\*\* is that any proprietary data you send us will be kept entirely private, heavily anonymised, processed in a secure way, and will never be sold to any commercial AI company.
Try the Model & Download the Weights:-
Official website or web interface:
https://www.arcleintelligence.com
\* \*\*Hugging Face (Weights & Model Card):\*\* \[Lucifer2006/Arcle-V1\](https://huggingface.co/Lucifer2006/Arcle-V1)
\* \*\*GitHub Repository:\*\* \[github.com/lucifertkod/personal-website\](https://github.com/lucifertkod/personal-website)
Keep up to date every day with news about Arcle V2: \[x.com/Anonomus090806\](https://x.com/Anonomus090806) | \[Instagram: @arcleintelligence.ai\](https://www.instagram.com/arcleintelligence.ai)
It was said that one boy from Bihar could not manage to create a real omni model, and an anonymous website asserted that I had built nothing whatsoever.
Let us jointly demonstrate that the open-source community is capable of building anything that a centralised corporation can.
Download it, run it, test it and break it, and then give us your feedback.
— \*\*Abhinav Anand\*\*
Founder, ArcleIntelligence
Source: r/AIxProduct · by /u/That-Bookkeeper-8316