ON
← Back to feed
United StatesBusiness2 days ago

A startup claims it broke through a bottleneck that’s holding back LLMs

A Miami-based AI startup named Subquadratic has claimed to have overcome a major mathematical bottleneck limiting large language models (LLMs). The company introduced a new model called SubQ, which it says is faster, cheaper, and more energy-efficient than existing models. SubQ is reported to handle significantly more text at once and perform well on key tasks compared to leading models from companies like Google DeepMind, OpenAI, and Anthropic. However, initial skepticism existed due to limited evidence, though the company has since shared results from an independent evaluation.

Subquadratic has now shared more details about its new model. But some are still skeptical.

June 19, 2026

Stephanie Arnett/MIT Technology Review | Adobe Stock

Miami-based AI startup Subquadratic came out of stealth mode last month with a huge claim. It announced that it had solved a mathematical bottleneck that had been holding back large language models for almost a decade.

The details were thin, and many people were unconvinced. But Subquadratic has started to bring the receipts, sharing the results of an independent evaluation of its new tech. The results suggest that the company’s claims might be worth paying attention to.

According to Subquadratic, it has developed a new kind of LLM, called SubQ, that is faster and cheaper and uses a lot less energy than any other model on the market. The company also claims that SubQ is able to process up to 12 times as much text at once than most other models, allowing it to carry out a range of data-heavy tasks, such as analyzing hundreds of documents or entire code bases.

What’s more, Subquadratic says, SubQ does this while more or less matching the performance of the best models put out by Google DeepMind, OpenAI, and Anthropic on key tasks like coding.

The problem was that the company at first provided little evidence for its claims beyond a handful of self-published test scores. And it has yet to make SubQ widely available for people to try out themselves.

So it’s no surprise that Subquadratic’s claims were met with skepticism. Dan McAteer, an artificial intelligence engineer, captured the overall response on X : “SubQ is either the biggest breakthrough since the Transformer ... or it’s AI Theranos.”

A month on, the company has published more information about its model , including the results of additional independent tests run by third-party firm Appen.

“We expected healthy skepticism,” says Subquadratic cofounder and chief technology officer Alex Whedon. “In hindsight, releasing the third-party benchmarks alongside the initial announcement would have preempted much of the skepticism, which is why we’re taking the time to make sure any future results are fully verified before putting them out.”

Subquadratic asked Appen, which evaluates other companies’ models, to run its tests on SubQ. The results seem to back up a lot of Subquadratic’s claims. “That was really exciting to me, it validated their architecture,” says Jeanine Sinanan-Singh, Appen’s director of generative AI research.

“I was like, ‘Wow, this could be a game changer,’ because models struggle with speed and inefficiency,” she adds. “But when you have kind of shocking results, it’s really not as credible when you say it yourself.”

SubQ won’t replace existing top models across the board, but it could offer huge increases in speed at a fraction of the typical cost for certain tasks. Subquadratic insists that in the long run, though, its breakthrough could change how LLMs are built. “We hope we’re kicking off a new age of efficiency,” says Justin Dangel, the firm’s cofounder and CEO. “We don’t think anybody will be building on transformers in a few years.”

Attention!

To understand why Subquadratic’s claims are a big deal, let’s dig into how most LLMs work. The key mechanism inside an LLM is a type of neural network called a transformer, which runs a process known as dense attention. Today’s LLMs typically chain together multiple transformers. (The foundational paper of the LLM era, published by researchers at Google in 2017, was titled “Attention Is All You Need.” )

Dense attention works like this: When a transformer processes a chunk of text , it first encodes each word (or part of a word, known as a token) with a number. To capture the meaning of the full text, it then multiplies each of those numbers with every other number for that text. For example, a piece of text 10,000 words long would kick off almost 50 million individual multiplications. That’s a lot of computation and the main reason that LLMs are notorious power hogs.

“If you want to summarize The Great Gatsby , you have to look at the first word and the last word together, and then you have to look at every other combination,” says Dangel.

As the length of the text increases, the number of computations skyrockets. That’s because each additional number must be multiplied by all other previous numbers. Double the number of words, and you roughly quadruple the number of computations, a rate of increase known as a quadratic expansion.

(You can picture this yourself: Draw a circle and mark dots around its edge. Each dot is a token. Then draw lines between pairs of dots to represent the multiplication of those two tokens. A circle with five dots will have 10 lines crossing it. Make it 10 dots and you will have 45 lines, 20 dots and you will have 190 lines, and so on.)

Slashing costs

Subquadratic’s solution is to ditch dense attention, the core operation of a transformer, in favor of what’s known as sparse attention, which slashes the numbe…

Read the full article at MIT Technology Review
Source document: Subquadratic's Independent Evaluation Results

3 reports

TechCrunchParty-alignedCenter2 days ago
The CEO of Allbirds’ new AI biz has a plan, but no employees

Allbirds, previously known for its direct-to-consumer shoe business, has shifted focus to AI through its newly rebranded entity Smartbird. The company sold its shoe division for $43 million and raised additional funds from the stock market. Nadia Carlsten, former AWS executive and CEO of Smartbird, is tasked with building a new team and establishing an office. The article notes that Smartbird currently operates as a 'startup with a sole founder and a very large seed round,' though its future plans remain unclear.

Bias read (Center): The article provides a factual overview of Allbirds' pivot to AI and Smartbird's current status without overtly favoring any particular perspective. It includes quotes from the CEO and contextualizes the company's actions within broader trends in Silicon Valley and the AI industry. There is no clear

MIT Technology ReviewIndependentCenter2 days ago
A startup claims it broke through a bottleneck that’s holding back LLMs

A Miami-based AI startup named Subquadratic has claimed to have overcome a major mathematical bottleneck limiting large language models (LLMs). The company introduced a new model called SubQ, which it says is faster, cheaper, and more energy-efficient than existing models. SubQ is reported to handle significantly more text at once and perform well on key tasks compared to leading models from companies like Google DeepMind, OpenAI, and Anthropic. However, initial skepticism existed due to limited evidence, though the company has since shared results from an independent evaluation.

Bias read (Center): The article presents information objectively without overtly favoring one side. It reports on the claims made by Subquadratic, mentions the initial skepticism, and notes the company's efforts to provide supporting evidence. There is no clear ideological framing or biased language.

Official sources cited

  • organisation Subquadratic's Independent Evaluation Results
SlateIndependentCenter2 days ago
Can ChatGPT Be a Criminal Accomplice?

The article discusses concerns about large language models (LLMs), such as ChatGPT, providing harmful advice despite supposed safeguards. It references real-world cases where LLM-generated content was used to plan a mass shooting. The episode features guest Mark Follman, who has written about preventing mass shootings.

Bias read (Center): The article presents a factual discussion on the potential misuse of AI technology without overtly favoring any political perspective. It highlights concerns raised by experts and includes a guest with relevant expertise, maintaining a balanced tone.

Official sources cited

  • press release Trigger Points: Inside the Mission to Stop Mass Shootings in America

Go to the primary sources (2)

The official sources this coverage is built on. Read them directly to bypass framing.

  • organisationSubquadratic's Independent Evaluation Results
  • press_releaseTrigger Points: Inside the Mission to Stop Mass Shootings in America