# Ilya Sutskever

> Source: https://aiwiki.ai/wiki/ilya_sutskever
> Updated: 2026-07-31
> Fact-checked: 2026-07-31
> Categories: AI Safety, Deep Learning, OpenAI, People
> License: CC BY 4.0 (https://creativecommons.org/licenses/by/4.0/) - attribute to "AI Wiki (aiwiki.ai)"
> Cite as: AI Wiki. "Ilya Sutskever." aiwiki.ai, 31 Jul 2026. https://aiwiki.ai/wiki/ilya_sutskever
> From AI Wiki (https://aiwiki.ai), the free encyclopedia of artificial intelligence. Reuse freely with attribution.

**Ilya Sutskever** (born 1986) is a Russian-born computer scientist who was raised in Israel and moved to Canada as a teenager. He is a co-founder and chief executive of [Safe Superintelligence Inc.](https://aiwiki.ai/wiki/safe_superintelligence), and was a co-founder, research director, and chief scientist of [OpenAI](https://aiwiki.ai/wiki/openai). Earlier, he studied with [Geoffrey Hinton](https://aiwiki.ai/wiki/geoffrey_hinton) at the University of Toronto and worked at Google Brain.

Sutskever's research record includes coauthoring the 2012 [AlexNet](https://aiwiki.ai/wiki/alexnet) paper, leading the 2014 sequence-to-sequence paper, and coauthoring the GPT-3 and GPT-4 reports. These works contributed to major changes in [computer vision](https://aiwiki.ai/wiki/computer_vision), machine translation, and large-scale language modeling. His public profile also reflects two later roles: advocating for research on the control of more capable AI systems, and participating in OpenAI's November 2023 board decision to remove Sam Altman. The published record supports a substantial leadership role, but products such as ChatGPT and broad research programs at Google and OpenAI were team efforts rather than the work of one person.

## Early life and education

Sutskever was born in 1986 in Nizhny Novgorod, Russia, and was raised mostly in Jerusalem. He recalled first seeing a computer at about age five and becoming interested in programming. His family moved to Canada when he was 16. The University of Toronto admitted him from Grade 11 into its mathematics program on the basis of coursework at the Open University of Israel, after which he entered second- and third-year university courses.[1]

He earned an honours bachelor's degree in mathematics in 2005 and a master's degree in computer science in 2007. The University of Toronto School of Graduate Studies lists his PhD in computer science as awarded in 2013.[2] His thesis, *Training Recurrent Neural Networks*, is also dated 2013 and describes methods for training [recurrent neural networks](https://aiwiki.ai/wiki/recurrent_neural_network) on long-term dependencies, text, and other sequential data.[3] The institutional graduate record and the thesis therefore support 2013 as the degree year, although a later university magazine profile gives 2012.

Sutskever's doctoral adviser was Hinton. After his doctoral work, Sutskever was a postdoctoral researcher in [Andrew Ng](https://aiwiki.ai/wiki/andrew_ng)'s group at [Stanford University](https://aiwiki.ai/wiki/stanford_university). His own professional biography then places DNNresearch and three years as a research scientist at Google Brain before OpenAI.[4]

## Research career

### AlexNet and DNNresearch

Krizhevsky, Sutskever, and Hinton incorporated DNNresearch in 2012. Google acquired the company in March 2013 for undisclosed terms. The University of Toronto announcement said Krizhevsky and Sutskever would move to Google, while Hinton would divide his time between the university and Google.[5]

In 2012, Sutskever, Alex Krizhevsky, and Hinton published "ImageNet Classification with Deep Convolutional Neural Networks." The model described in the paper had 60 million parameters and 650,000 neurons, with five convolutional layers followed by three fully connected layers. It was trained for five to six days on two NVIDIA GTX 580 [GPUs](https://aiwiki.ai/wiki/gpu). On the [ImageNet](https://aiwiki.ai/wiki/imagenet) test set, the authors reported a 15.3% top-5 error rate, compared with 26.2% for the second-best entry, a difference of 10.9 percentage points.[6]

The result showed that a large [convolutional neural network](https://aiwiki.ai/wiki/convolutional_neural_network), trained with enough labeled data and graphics processing, could outperform the leading competition systems by a large margin. It became an important demonstration of the practical value of [deep learning](https://aiwiki.ai/wiki/deep_learning), while the implementation and experimental result were the work of all three authors rather than evidence of sole credit for any one contributor.

### Google Brain, sequence-to-sequence learning, and other work

At [Google Brain](https://aiwiki.ai/wiki/google_brain), Sutskever worked on neural approaches to language and other large-scale learning problems. His most influential paper from this period was "Sequence to Sequence Learning with Neural Networks," written with Oriol Vinyals and Quoc V. Le. The method used one multilayer LSTM to encode a variable-length input sequence into a fixed-dimensional representation and a second LSTM to generate the output sequence.[7]

On English-to-French translation, the paper reported a [BLEU](https://aiwiki.ai/wiki/bleu_bilingual_evaluation_understudy) score of 34.8 for direct translation. Using the model to rerank 1,000 hypotheses from another system produced 36.5. Reversing the source sentence improved optimization by shortening important dependencies between source and target words. The work helped establish an end-to-end encoder-decoder approach to [machine translation](https://aiwiki.ai/wiki/machine_translation) and other variable-length sequence tasks.

Sutskever was also a coauthor of the 2016 Nature paper on [AlphaGo](https://aiwiki.ai/wiki/alphago), the system that combined policy and value networks with tree search to play Go.[8] This establishes his authorship on the published research, but it does not by itself identify which components he personally designed.

## OpenAI

### Founding and research leadership

OpenAI was announced in December 2015 as a nonprofit artificial-intelligence research company. Its launch statement named Sutskever as research director and Greg Brockman as chief technology officer. It also said that funders had committed $1 billion, while expecting the organization to spend only a small fraction of that amount in its first several years.[9] The commitment should not be described as an acquisition price or as $1 billion already spent.

Sutskever later served as chief scientist. His role placed him in the research leadership of the organization as it increasingly trained large [language models](https://aiwiki.ai/wiki/large_language_model). He was a coauthor of the GPT-3 paper, which described a 175-billion-parameter autoregressive model evaluated through zero-shot, one-shot, and few-shot prompting.[10] He was also among the authors of the GPT-4 technical report. That report deliberately withheld architecture, model size, hardware, training compute, dataset construction, and related implementation details, so it cannot support more specific claims about his personal technical contribution.[11]

OpenAI's GPT work formed the technical basis for later products, including [ChatGPT](https://aiwiki.ai/wiki/chatgpt). Sutskever's leadership and paper authorship make him part of that history, but public sources do not isolate individual responsibility for ChatGPT's product design, data pipeline, post-training, or launch. It is therefore more accurate to describe his role in the research organization and the papers than to credit him personally with building the product.

Sutskever is often associated with the view that larger models, datasets, and training budgets would continue to improve capability. The success of AlexNet, sequence-to-sequence learning, and the GPT program is consistent with that emphasis. However, [scaling laws](https://aiwiki.ai/wiki/scaling_laws) were studied by many researchers, and the evidence does not support treating the "scaling hypothesis" as a theory invented by Sutskever alone.

### Superalignment

In July 2023, OpenAI created a Superalignment team co-led by Sutskever and Jan Leike. The announcement set a four-year goal of developing methods to steer and control systems much more capable than humans and said OpenAI would dedicate 20% of the compute it had secured to date to the effort. Its proposed research directions included scalable training, validation, stress testing, interpretability, and adversarial evaluation.[12]

The wording of that announcement matters. It was a research goal and a planned allocation, not a claim that alignment had been solved or that 20% of all future OpenAI compute was guaranteed.

### November 2023 board crisis

On November 17, 2023, OpenAI's nonprofit board removed [Sam Altman](https://aiwiki.ai/wiki/sam_altman) as chief executive. The company's announcement said a board review had concluded that Altman was not consistently candid in communications with the board and that the board no longer had confidence in his leadership. Sutskever was then both a board member and chief scientist.[13]

OpenAI's March 2024 summary of a WilmerHale review gave a narrower account than much contemporary speculation. According to the company, the review found a breakdown in trust between the prior board and Altman. It said the decision did not arise from concerns about product safety or security, the pace of development, OpenAI's finances, or statements to investors and customers. It also said the board acted within its discretion, but on an abridged timetable, without a full inquiry or an opportunity for Altman to answer its concerns, and that his conduct did not mandate removal.[14] The company published a summary rather than the full investigation, so this is OpenAI's account of the review, not an independently inspectable final report.

The decision triggered a staff revolt. Axios reported that more than 500 of OpenAI's 700-plus employees threatened to resign unless the board stepped down. On November 20, Sutskever publicly said he deeply regretted participating in the board's actions and wanted to reunite the company.[15] Altman returned within days under a reconstituted board, on which Sutskever did not serve.

### Departure and later legal record

Sutskever announced his departure from OpenAI on May 14, 2024. The Associated Press reported that Jakub Pachocki replaced him as chief scientist and that Sutskever said he was working on a personally meaningful new project.[16] SSI was announced the following month.

Additional details about the 2023 crisis entered the public record through litigation between Elon Musk and OpenAI. Selected excerpts from Sutskever's October 1, 2025 deposition show that he sent a 52-page memo to the three independent directors and believed termination was appropriate. In the excerpts, he said many or all screenshots came from Mira Murati. For specific allegations about conversations and executive departures, he acknowledged that Murati was his source and that he had not independently verified the information with the people involved. The excerpts also describe a brief post-removal proposal involving a merger with [Anthropic](https://aiwiki.ai/wiki/anthropic); Sutskever said he opposed it, while recalling that other directors were more supportive.[17]

These statements require careful treatment. The publicly archived document contains selected pages rather than the complete deposition, and the memo's allegations are not judicial findings of fact. A later federal discovery order confirms disputes over a separate Brockman memo and the value of Sutskever's financial interest in OpenAI, and it authorized a second deposition on those subjects.[18]

At trial in May 2026, Reuters reported that Sutskever testified he had considered taking action to remove Altman for at least a year before the November 2023 vote. The same report said he valued his ownership stake in OpenAI at about $7 billion.[19] That testimony supports the approximate dollar value he gave at trial, but not an inferred ownership percentage or a precise company valuation.

## Safe Superintelligence Inc.

### Founding and stated mission

Sutskever founded SSI in June 2024 with [Daniel Gross](https://aiwiki.ai/wiki/daniel_gross) and [Daniel Levy](https://aiwiki.ai/wiki/daniel_levy). The company describes itself as an American laboratory with offices in Palo Alto and Tel Aviv, one goal, and one product: a safe superintelligence. Its launch statement says it intends to advance capability and safety together and to insulate the work from short-term commercial pressure and product cycles.[20]

These are the founders' stated objectives, not evidence that a safe superintelligent system already exists or that SSI's approach has been independently validated. As of the research cutoff, SSI had not publicly released a model, product, evaluation, architecture, or detailed safety method.

### Funding, leadership, and compute

SSI's own updates record three milestones. It announced a $1 billion raise on September 4, 2024. It said Gross left on June 29, 2025, after which Sutskever became chief executive, Levy became president, and the technical team continued to report to Sutskever. On July 26, 2026, SSI said a partnership with NVIDIA would increase its compute tenfold.[21]

Public reporting provides additional financial context:

| Date | Publicly reported event | Amount and valuation | Evidence status |
| --- | --- | --- | --- |
| September 2024 | Initial external financing | $1 billion raised; Reuters sources reported a $5 billion valuation | Company confirmation of amount; valuation attributed to Reuters sources.[22] |
| April 2025 | Additional financing | TechCrunch, citing the Financial Times, reported $2 billion at a $32 billion valuation | Reported terms; SSI did not comment.[23] |

The first two reported rounds total $3 billion. An unsupported higher cumulative figure previously stated for the period before the NVIDIA agreement is removed.

Reuters reported in April 2025 that Alphabet and NVIDIA had joined SSI's investors, while the exact terms were not determined. The report said Google would sell SSI access to significant quantities of its custom Tensor Processing Units (TPUs) and that SSI was then using TPUs primarily rather than GPUs.[24] Google later named Safe Superintelligence among research labs using TPUs in its July 2025 earnings remarks.[25] Neither source establishes a superlative ranking among Google's external TPU customers.

### Publicly described research direction

SSI has disclosed little technical detail, so Sutskever's public talks are the main evidence about his research outlook rather than a specification of the company's system. At NeurIPS in December 2024, he argued that internet-scale pretraining would eventually end because compute could grow while the supply of internet data could not. He suggested synthetic data, answer selection, and real-world data as possible directions, and predicted that systems with stronger reasoning would become less predictable. These were forecasts and research opinions, not demonstrated properties of SSI technology.[26]

In a November 2025 interview with Dwarkesh Patel, Sutskever described roughly 2012 to 2020 as an "age of research," 2020 to 2025 as an "age of scaling," and the next period as a return to research. He argued that current systems show a gap between benchmark performance and reliable generalization, and said SSI was investigating ideas about human-like continual learning. When asked for a timeline to a system that learns as well as a human and could then become superhuman, he gave a range of 5 to 20 years. That range was explicitly his forecast, not a measured development schedule.[27]

On July 27, 2026, SSI and [NVIDIA](https://aiwiki.ai/wiki/nvidia) jointly announced a long-term partnership. The release said SSI would receive access to the NVIDIA Vera Rubin platform, increase its compute by an order of magnitude, and collaborate on future compute platforms. NVIDIA made an investment but did not disclose its value in the announcement. The release also attributed to SSI the claim that it had pursued a new research direction for two years and was ready to scale it.[28] Reuters separately reported, citing a person briefed on the deal, that the investment was $5 billion in equity.[29] The official disclosure and the attributed report should remain distinct.

## Research views and influence

Sutskever's career connects three recurring ideas. First, learned representations can replace handcrafted task-specific features when models receive enough data and compute. AlexNet was a prominent result for images, while sequence-to-sequence learning applied an end-to-end learned representation to language. Second, increasing model and dataset scale can yield capabilities not obvious from smaller experiments. His OpenAI leadership coincided with a concentrated bet on this approach. Third, stronger systems create a harder control problem, which motivated his Superalignment work and SSI's stated mission.

His later comments do not amount to a rejection of all scaling. In the 2025 interview, he said additional scale could still improve current approaches, but argued that it was unlikely by itself to produce the reliable generalization and continual learning he sought.[27] The distinction is between ending progress and ending a particular source of relatively predictable progress from more internet pretraining data.

Sutskever has described his research judgment as guided by simple, brain-inspired principles and by a preference for ideas that appear coherent from multiple angles. This is a personal account of method rather than a reproducible scientific criterion. His most verifiable influence remains the published work, the teams he led, and the research directions he publicly advocated.

## Awards and recognition

| Year | Recognition | Basis |
| --- | --- | --- |
| 2022 | NeurIPS Test of Time Award | Coauthor of the 2012 AlexNet paper.[30] |
| 2022 | Fellow of the Royal Society | Elected for contributions to advances from object recognition to natural-language processing.[31] |
| 2023 | NeurIPS Test of Time Award | Coauthor of "Distributed Representations of Words and Phrases and their Compositionality."[32] |
| 2024 | NeurIPS Test of Time Award | Lead author of "Sequence to Sequence Learning with Neural Networks."[33] |
| 2024 | TIME100 AI | Included in TIME's list of influential people in artificial intelligence, not the general TIME100 list.[34] |
| 2025 | University of Toronto honorary Doctor of Science | Conferred June 6, 2025 for his research and work on responsible AI.[35] |
| 2026 | NAS Award for the Industrial Application of Science | The National Academy of Sciences said it was the award's first presentation for artificial intelligence; the prize is $25,000.[36] |

Sutskever was also one of the authors of the 2016 systems paper describing [TensorFlow](https://aiwiki.ai/wiki/tensorflow).[37] As with AlphaGo and the GPT reports, authorship is the supported biographical fact; a more specific allocation of implementation credit would require additional evidence.

## See also

- [AI alignment](https://aiwiki.ai/wiki/ai_alignment)
- [Artificial general intelligence](https://aiwiki.ai/wiki/artificial_general_intelligence)
- [Backpropagation](https://aiwiki.ai/wiki/backpropagation)
- [Natural language processing](https://aiwiki.ai/wiki/natural_language_processing)
- [Neural machine translation](https://aiwiki.ai/wiki/neural_machine_translation)
- [Superintelligence](https://aiwiki.ai/wiki/superintelligence)

## References

[1] Alec Scott. "Heard It Through the AI." *University of Toronto Magazine*, September 28, 2022. https://magazine.utoronto.ca/people/alumni-donors/heard-it-through-the-ai-openai-ilya-sutskever/

[2] "Ilya Sutskever." *University of Toronto School of Graduate Studies*. https://www.sgs.utoronto.ca/profile/ilya-sutskever/

[3] Ilya Sutskever. *Training Recurrent Neural Networks*. PhD thesis, University of Toronto, 2013. https://www.cs.toronto.edu/~ilya/pubs/ilya_sutskever_phd_thesis.pdf

[4] Ilya Sutskever. "Ilya Sutskever's home page." *University of Toronto Department of Computer Science*. https://www.cs.toronto.edu/~ilya/

[5] Sara Franca. "Google acquires U of T neural networks company." *University of Toronto*, March 12, 2013. https://www.utoronto.ca/news/google-acquires-u-t-neural-networks-company

[6] Alex Krizhevsky, Ilya Sutskever, and Geoffrey E. Hinton. "ImageNet Classification with Deep Convolutional Neural Networks." *NeurIPS 2012*. https://proceedings.neurips.cc/paper/2012/hash/c399862d3b9d6b76c8436e924a68c45b-Abstract.html

[7] Ilya Sutskever, Oriol Vinyals, and Quoc V. Le. "Sequence to Sequence Learning with Neural Networks." *NeurIPS 2014*. https://proceedings.neurips.cc/paper_files/paper/2014/hash/5a18e133cbf9f257297f410bb7eca942-Abstract.html

[8] David Silver et al. "Mastering the game of Go with deep neural networks and tree search." *Nature*, January 27, 2016. https://www.nature.com/articles/nature16961

[9] OpenAI. "Introducing OpenAI." December 11, 2015. https://openai.com/index/introducing-openai/

[10] Tom B. Brown et al. "Language Models are Few-Shot Learners." *arXiv*, May 28, 2020. https://arxiv.org/abs/2005.14165

[11] OpenAI. "GPT-4 Technical Report." *arXiv*, March 15, 2023. https://arxiv.org/abs/2303.08774

[12] Jan Leike and Ilya Sutskever. "Introducing superalignment." *OpenAI*, July 5, 2023. https://openai.com/index/introducing-superalignment/

[13] OpenAI. "OpenAI announces leadership transition." November 17, 2023. https://openai.com/index/openai-announces-leadership-transition/

[14] OpenAI. "Review completed and Altman and Brockman to continue to lead OpenAI." March 8, 2024. https://openai.com/index/review-completed-altman-brockman-to-continue-to-lead-openai/

[15] Ina Fried. "OpenAI's Sutskever says he regrets board's firing of Altman." *Axios*, November 20, 2023. https://www.axios.com/2023/11/20/sam-altman-fired-openai-board-illya-sutsever-regrets

[16] Associated Press. "OpenAI co-founder Ilya Sutskever announces departure from ChatGPT maker." May 14, 2024. https://finance.yahoo.com/news/openai-co-founder-ilya-sutskever-002159131.html

[17] "Selected excerpts from the deposition of Ilya Sutskever." *Musk v. Altman*, October 1, 2025. https://chatgptiseatingtheworld.com/wp-content/uploads/2025/11/Ilya-Sutsekever-Deposition.df_.pdf

[18] Thomas S. Hixson. "Discovery Order re 301 Discovery Letter Brief." *United States District Court for the Northern District of California*, October 17, 2025. https://docs.justia.com/cases/federal/district-courts/california/candce/4%3A2024cv04722/433688/324

[19] Reuters. "Former OpenAI executive Sutskever discloses nearly $7 billion stake in AI firm." May 11, 2026. https://m.investing.com/news/stock-market-news/former-openai-executive-sutskever-discloses-nearly-7-billion-stake-in-ai-firm-4677950?ampMode=1

[20] Safe Superintelligence Inc. "Safe Superintelligence Inc." June 19, 2024. https://ssi.inc/

[21] Safe Superintelligence Inc. "Updates." Through July 26, 2026. https://ssi.inc/updates

[22] Kenrick Cai, Krystal Hu, and Anna Tong. "OpenAI co-founder Sutskever's new safety-focused AI startup SSI raises $1 billion." *Reuters*, September 4, 2024. https://www.marketscreener.com/quote/stock/MICROSOFT-CORPORATION-4835/news/Exclusive-OpenAI-co-founder-Sutskever-s-new-safety-focused-AI-startup-SSI-raises-1-billion-47802334/

[23] Anthony Ha. "OpenAI co-founder Ilya Sutskever's Safe Superintelligence reportedly valued at $32B." *TechCrunch*, April 12, 2025. https://techcrunch.com/2025/04/12/openai-co-founder-ilya-sutskevers-safe-superintelligence-reportedly-valued-at-32b/

[24] Kenrick Cai and Krystal Hu. "Alphabet, Nvidia invest in OpenAI co-founder Sutskever's SSI." *Reuters*, April 12, 2025. https://m.au.investing.com/news/stock-market-news/exclusivealphabet-nvidia-invest-in-openai-cofounder-sutskevers-ssi-source-says-3781978?ampMode=1

[25] Sundar Pichai. "Alphabet Q2 2025 earnings remarks." *Google*, July 23, 2025. https://blog.google/company-news/inside-google/message-ceo/alphabet-earnings-q2-2025/

[26] Reuters. "AI with reasoning power will be less predictable, Ilya Sutskever says." December 14, 2024. https://indianexpress.com/article/technology/artificial-intelligence/ai-with-reasoning-power-will-be-less-predictable-ilya-sutskever-9724266/

[27] Dwarkesh Patel. "We're moving from the age of scaling to the age of research." *Dwarkesh Podcast*, November 25, 2025. https://www.dwarkesh.com/p/ilya-sutskever-2

[28] NVIDIA. "Ilya Sutskever's Safe Superintelligence Inc. and NVIDIA announce long-term strategic partnership." July 27, 2026. https://nvidianews.nvidia.com/news/ilya-sutskevers-safe-superintelligence-inc-and-nvidia-announce-long-term-strategic-partnership

[29] Deepa Seetharaman. "Nvidia to invest $5 billion in Ilya Sutskever's AI startup, source says." *Reuters*, July 27, 2026. https://ca.marketscreener.com/news/nvidia-to-invest-5-billion-in-ilya-sutskever-s-ai-startup-source-says-ce7f51dcd18ff022

[30] NeurIPS. "Announcing the NeurIPS 2022 Awards." November 21, 2022. https://blog.neurips.cc/2022/11/21/announcing-the-neurips-2022-awards/

[31] Royal Society. "Outstanding scientists elected as Fellows and Foreign Members of the Royal Society." May 10, 2022. https://royalsociety.org/news/2022/05/new-fellows-2022/

[32] NeurIPS. "Announcing the NeurIPS 2023 Paper Awards." December 11, 2023. https://blog.neurips.cc/2023/12/11/announcing-the-neurips-2023-paper-awards/

[33] NeurIPS. "Announcing the NeurIPS 2024 Test of Time Paper Awards." November 27, 2024. https://blog.neurips.cc/2024/11/27/announcing-the-neurips-2024-test-of-time-paper-awards/

[34] Billy Perrigo. "Ilya Sutskever: The 100 Most Influential People in AI 2024." *TIME*, September 5, 2024. https://time.com/collections/time100-ai-2024/7012869/ilya-sutskever-2/

[35] Rahul Kalvapalle. "Ilya Sutskever, a leader in AI and its responsible development, receives U of T honorary degree." *University of Toronto*, June 6, 2025. https://www.utoronto.ca/news/ilya-sutskever-leader-ai-and-its-responsible-development-receives-u-t-honorary-degree

[36] National Academy of Sciences. "NAS Award for the Industrial Application of Science." 2026. https://www.nasonline.org/award/nas-award-for-the-industrial-application-of-science/

[37] Martín Abadi et al. "TensorFlow: Large-Scale Machine Learning on Heterogeneous Distributed Systems." *arXiv*, March 14, 2016. https://arxiv.org/abs/1603.04467

