AI Image Generation
AI image generation is the use of artificial intelligence systems to create visual content, including photographs, illustrations, paintings, concept art, and graphic designs, from text descriptions, reference…
Explore Computer Vision through related topics and the articles other pages reference most.
Articles that also belong to these categories. Counts cover all of Computer Vision.
Showing 1-13 of 13 articles
AI image generation is the use of artificial intelligence systems to create visual content, including photographs, illustrations, paintings, concept art, and graphic designs, from text descriptions, reference…
AI video generation is the use of artificial intelligence systems, predominantly diffusion transformers, to create video clips from text descriptions, still images, or other video, producing sequences of…
AI in agriculture refers to the application of artificial intelligence, machine learning, computer vision, and robotics to farming and food production.
Autonomous driving is the sustained performance by a vehicle system of part or all of the dynamic driving task.
An autonomous vehicle, also known as a self-driving car, driverless car, or robotic vehicle, is a vehicle that uses artificial intelligence, sensors, and software to navigate and operate without human input.
Computer vision is the study of computational methods that extract, estimate, or generate useful representations from visual measurements.
A computer-use agent (CUA) is a category of AI agent in artificial intelligence that performs tasks by directly operating a general-purpose computer's graphical user interface (GUI) the way a human does, by…
A deepfake is synthetic media in which a real person's face, voice, or body is digitally replaced, manipulated, or fabricated using artificial intelligence, most often deep learning techniques such as…
Fei-Fei Li (born 1976) is a Chinese-American computer scientist, the inaugural Sequoia Capital Professor of Computer Science at Stanford University, and co-director of the Stanford Institute for Human-Centered…
LeNet is the pioneering family of convolutional neural networks developed by Yann LeCun and collaborators at AT&T Bell Labs between roughly 1988 and 1998 to read handwritten characters
OCR Models are artificial intelligence (AI) systems that convert images of typed, handwritten, or printed text into machine-readable digital text through Optical Character Recognition (OCR).
Object detection is a computer vision task that finds instances of interest in an image and assigns each one a category.
Pre-training is a stage of machine learning in which a model learns parameters from a source dataset or source objective before those parameters are reused or adapted for a target use.