# How AI Generates Images Step by Step

AI image generators do not retrieve or collage existing photos; instead, they translate text into mathematical coordinates to guide a process that progressively removes static noise from a blank, compressed digital canvas. This complex pipeline relies on diffusion models to organize random pixels into recognizable patterns based on learned statistical probabilities. The result is an entirely novel image sculpted from mathematical chaos, operating in a way that mimics thermodynamics rather than a traditional search engine.

## The Collage Myth Versus Statistical Reality

Before exploring the technical architecture of how modern artificial intelligence synthesizes visuals, it is crucial to dismantle the most pervasive misconception surrounding the technology: the "collage" theory. 

A widespread belief among the general public and traditional artists is that generative AI acts as a highly sophisticated search engine [cite: 1, 2]. According to this myth, when a user requests an image, the AI scours a massive internal database of copyrighted artwork, cuts these existing images into digital puzzle pieces, and seamlessly stitches them together to form a "new" composition [cite: 1, 2]. Because human digital artists often use photo-bashing and collage techniques, assuming the machine operates the same way is a rational, albeit mathematically incorrect, assumption [cite: 1].

In reality, generative AI models do not store a database of images inside their underlying code [cite: 3]. When a model is trained on a dataset—such as the massive LAION-5B database, which contains billions of image-text pairs—it is not saving those files to a hard drive [cite: 3, 4]. Instead, it is analyzing them to learn the *statistical patterns* and spatial relationships of visual data [cite: 3, 5]. 

If an AI analyzes a million photographs of a dog, it learns the mathematical probability of where a wet nose, specific fur textures, and floppy ears should appear in relation to lighting and background environments. Once the training is complete, the original images are discarded [cite: 3]. When you prompt the AI to generate a dog, it does not retrieve a stored photo. It draws upon its learned statistical weights to organize completely random pixels into a coherent pattern from scratch [cite: 3]. The machine is generating entirely new pixel arrangements based on mathematical probabilities, a process that requires a multi-step pipeline to translate human language into a high-fidelity visual output [cite: 6].

[image delta #1, 0 bytes]





## Step 1: Text Encoding (Bridging Language and Mathematics)

If a computer only understands numerical arrays, how does it comprehend a prompt like "a cyberpunk golden retriever drinking coffee in neon rain"? The critical bridge between human linguistics and machine vision is the text encoder. When a user submits a prompt, the system first breaks the sentence down into manageable chunks called "tokens." These tokens are then passed through an encoder, which translates the words into a multidimensional numerical vector known as an embedding [cite: 6, 7].

An embedding acts as a set of GPS coordinates for meaning within a vast, shared mathematical space [cite: 8]. In this high-dimensional architecture, concepts that share semantic meaning reside close to one another [cite: 6, 8]. For example, the vector for "puppy" and the vector for "dog" are located in the same mathematical neighborhood, while the vector for "galaxy" is positioned entirely elsewhere [cite: 6]. This spatial mapping allows the AI to understand relationships between words.

### The Foundation: Contrastive Language-Image Pretraining (CLIP)

The breakthrough that made modern text-to-image generation possible was the development of CLIP (Contrastive Language-Image Pretraining) by OpenAI in 2021 [cite: 9]. CLIP was designed to teach artificial intelligence how to connect visual content with descriptive text seamlessly [cite: 8, 10]. 

To train CLIP, researchers utilized 400 million pairs of images and their corresponding text captions scraped from the internet [cite: 9, 11]. The system employed two separate translating mechanisms: an image encoder (typically a ResNet50 or Vision Transformer model) and a text encoder [cite: 8, 11]. During training, the image encoder would scan a photo of a dog and output a numerical vector, while the text encoder would read the caption "a photo of a dog" and output its own vector [cite: 8, 9]. Crucially, both vectors were projected into the exact same shared dimensional space [cite: 8, 9].

The training objective utilized a method called contrastive learning [cite: 8, 11]. When given a batch of $N$ images and $N$ texts, the model computed a similarity score for every possible combination, resulting in an $N \times N$ grid [cite: 8, 9]. The model adjusted its internal weights to maximize the cosine similarity for the correct pairs (the diagonal of the grid) and minimize the similarity for incorrect pairings [cite: 8, 9, 11]. Through this rigorous mathematical matchmaking, CLIP learned an incredibly robust universal language of numbers where an image and its accurate description pointed in the exact same direction [cite: 8].

While revolutionary, relying strictly on CLIP introduced specific failure modes. CLIP behaves like a "bag of words" model; it excels at understanding the overall visual vibe or semantic association of a prompt but struggles profoundly with syntax, grammar, and character sequences [cite: 12, 13]. If you ask a CLIP-based model for "a red sphere on top of a blue cube," it understands "red," "blue," "sphere," and "cube," but often misattributes the colors or ignores the spatial command "on top of" [cite: 12]. This phenomenon, known as misattribution or concept bleeding, was a hallmark of early AI image generators [cite: 12]. Furthermore, because CLIP does not read words as sequential letters but as holistic concepts, it was historically incapable of spelling text correctly within an image [cite: 13].

### The Shift to T5 and Dual-Encoder Architectures

To resolve the limitations of CLIP, the developers of next-generation models like Stable Diffusion 3 (SD3) and FLUX.1 introduced a hybrid approach, augmenting the pipeline with massive text-to-text Large Language Models (LLMs) like T5-XXL (Text-to-Text Transfer Transformer) [cite: 13, 14, 15]. 

Unlike CLIP, which is purely contrastive, T5 is a generative language model trained on massive text corpora (like the C4 dataset) to understand deep linguistic structures, grammar, and sequential logic [cite: 13, 15, 16]. T5 actually "reads" the sequence of characters. When you request a neon sign that says "OPEN," T5 understands that O-P-E-N is a specific sequence of characters that must appear in order [cite: 13]. 

Modern models leverage both encoders simultaneously [cite: 14]. In FLUX.1, for instance, the CLIP encoder processes the first 77 tokens of a prompt to establish the overarching visual aesthetics and subject matter [cite: 14]. Concurrently, the T5-XXL encoder processes up to 512 tokens, extracting granular details, deep contextual relationships, and precise typography instructions [cite: 14, 17]. This dual-encoder strategy is the primary reason why contemporary models can follow complex, multi-paragraph prompts and flawlessly render legible text on storefronts or book covers [cite: 14, 17].

| Feature | CLIP Text Encoder | T5-XXL Text Encoder |
| :--- | :--- | :--- |
| **Primary Function** | Maps images and text into a shared multi-modal semantic space using contrastive learning. | Deep natural language understanding, grammar processing, and sequential text interpretation. |
| **Architectural Strength** | Exceptional at capturing the overall visual "vibe," artistic style, and broad subject matter. | Exceptional at understanding complex syntax, spatial relationships, and exact spelling. |
| **Architectural Weakness** | Suffers from concept bleeding; fails at spelling, character sequences, and complex spatial logic (e.g., "left of"). | Lacks direct visual training; can allocate parameter space to non-visual concepts, increasing computational overhead. |
| **Token Processing Limit** | Typically capped at 77 tokens due to its architectural design. | Can process up to 512 tokens, enabling paragraph-length descriptive prompts. |
| **Role in Modern Models** | Used for foundational visual-semantic alignment and aesthetic direction. | Used for granular prompt comprehension, multi-subject separation, and in-image text generation. |

*Table 1: A comparative analysis of the primary text encoders driving modern AI image generation pipelines [cite: 12, 13, 14, 15, 16].*

## Step 2: The Latent Space Canvas

Once the text prompt has been encoded into a guiding mathematical vector, the AI requires a canvas to generate the image. However, working directly with raw pixels is computationally devastating. A standard 512x512 RGB image contains 786,432 individual data points (dimensions) [cite: 18, 19]. Trying to train a deep neural network to calculate and manipulate nearly a million pixels simultaneously at every step of generation requires an unfeasible amount of VRAM and processing power [cite: 18, 19]. 

To bypass this hardware bottleneck, modern systems employ Latent Diffusion Models (LDMs), popularized by the original release of Stable Diffusion [cite: 20]. LDMs do not operate in "pixel space"; they operate in "latent space" [cite: 18]. 

### Compression and the Variational Autoencoder (VAE)

Latent space is a highly compressed, lower-dimensional mathematical representation of data [cite: 18, 19]. Before the diffusion process can occur, a separate neural network known as a Variational Autoencoder (VAE) is utilized [cite: 18]. The VAE consists of an encoder that compresses raw, high-resolution pixel data into a dense, compact latent vector, and a decoder that reconstructs the original image from this compressed state [cite: 18, 19]. 

For example, the VAE might compress a 786,432-dimension pixel image into a 64x64x4 latent grid, reducing the data footprint to just 16,384 dimensions [cite: 18, 19]. This massive reduction makes subsequent mathematical operations exponentially more efficient, allowing these complex models to run on consumer-grade GPUs rather than requiring enterprise-level supercomputers [cite: 18].

### The Geography of Latent Space

Latent space is not merely a zipped file format; it is a continuous, semantic geography [cite: 19, 21]. When the VAE compresses an image, it discards irrelevant pixel noise and retains the core semantic features [cite: 19]. In this lower-dimensional vector space, every point maps to a meaningful encoding of data [cite: 19]. 

Because the space is continuous and interpolative, moving slightly along a specific mathematical axis changes the corresponding image in a logical, semantic way [cite: 19, 21]. For instance, one direction in the latent space might correspond to the concept of "age," while another corresponds to "a smile." By performing basic vector arithmetic—such as taking a latent vector of a face and moving it along the mathematical axis corresponding to glasses—the model edits the underlying concept without having to manually adjust tens of thousands of individual pixels [cite: 19, 22]. 

When a user initiates an image generation request, the model does not start with a blank pixel canvas. It generates a small, mathematically dense matrix of pure random Gaussian noise directly inside this latent space [cite: 6, 18]. This invisible block of mathematical static serves as the raw material from which the final image will be carved [cite: 6, 23]. 

## Step 3: The Diffusion Engine

With the text embeddings established as a guide and the latent noise generated as a canvas, the core engine of the system takes over: the diffusion model. 

Introduced as a generative framework in 2015, diffusion models draw their conceptual inspiration from non-equilibrium thermodynamics and the physical process of diffusion, where particles spread from high concentration to low concentration [cite: 20, 24]. The models operate on a two-stage probabilistic framework that trains artificial intelligence to master the act of destruction in order to learn the act of creation [cite: 25].

[image delta #2, 0 bytes]



### Forward Diffusion: Controlled Destruction

The foundation of the model's knowledge is built during the training phase through a process called Forward Diffusion [cite: 7, 26]. The forward process is a fixed, predefined Markov chain that takes a clean, real image from the training dataset and systematically destroys it [cite: 20, 24, 25]. 

Over a series of hundreds or thousands of timesteps ($T$), the algorithm incrementally adds a small, calculated amount of random Gaussian noise to the image [cite: 20, 26]. Think of it like taking a pristine photograph and slowly sprinkling layers of sand over it [cite: 26]. The degradation is slow and smooth, governed by a mathematical noise schedule (often a linear or cosine schedule) that dictates exactly how much noise is applied at each step [cite: 26]. By the end of timestep $T$, the original structured data is entirely obliterated, leaving behind a matrix of near-pure, isotropic Gaussian noise [cite: 24, 25, 26]. 

This forward process is not a neural network learning anything; it is a rigid mathematical formulation [cite: 25]. However, it serves a vital purpose: it maps the exact trajectory from a structured image distribution to a chaotic noise distribution, providing the training data necessary for the next phase [cite: 20].

### Reverse Diffusion: The Denoising Loop

The true intelligence of the generative model lies in the Reverse Diffusion process [cite: 24, 25]. The goal here is to train a deep neural network to iteratively reverse the forward process—to look at a state of chaos and reconstruct order [cite: 20, 24].

Starting at timestep $T$ with a completely noisy latent image, the neural network acts as a noise estimator (or score function) [cite: 20, 25]. At each step, it analyzes the current noisy image and predicts exactly which mathematical components represent the added Gaussian noise [cite: 20, 26]. It then subtracts that estimated noise, stepping backward to timestep $T-1$ [cite: 20, 24]. 



The network does not attempt to recreate the final image in a single, impossible leap [cite: 26]. It takes dozens of small, iterative steps. With each pass, the noise is reduced, and structure begins to emerge from the static—first vague outlines, then broad colors, and finally, hyper-realistic fine details [cite: 25, 26]. 

When you use an AI image generator, the model is exclusively running this reverse process [cite: 6]. It initializes a brand new block of random noise and utilizes text conditioning via cross-attention mechanisms [cite: 7]. The text embeddings generated in Step 1 are injected into the neural network, acting as a gravitational pull that forces the denoising process to carve out a specific shape (e.g., a futuristic city) rather than returning a random unconditioned image [cite: 7, 24, 26].

## Steering the Output: Seeds, Steps, and Guidance Scales

To grant users control over this highly probabilistic mathematical system, AI generators expose several core parameters that dictate how the denoising algorithm behaves. Understanding these parameters is essential for advanced prompt engineering.

*   **The Seed:** The seed is a specific integer used to initialize the exact pattern of random noise in the latent space at the very beginning of the generation process [cite: 27, 28]. Because computers generate "randomness" algorithmically, providing the same seed number will always produce the exact same starting static [cite: 29]. If a user inputs the same text prompt, the same guidance parameters, and the exact same seed, the diffusion model will trace an identical mathematical path and produce an identical image [cite: 27, 28]. By locking the seed, users can tweak a single word in their prompt to see exactly how the model alters the composition, making the seed invaluable for character consistency and reproducible experimentation [cite: 27].
*   **Inference Steps:** This parameter dictates the number of iterative denoising loops the model will undergo [cite: 27, 29]. Setting the steps too low (e.g., under 15) usually results in an image that looks blurry, abstract, or structurally deformed, as the model has not had enough iterations to fully remove the noise [cite: 29, 30]. Increasing the steps refines the detail and crispness of the image [cite: 29]. However, this operates on a curve of diminishing returns; once the underlying noise is resolved, additional steps merely consume computing power without noticeably improving image quality, and in some cases, can over-process fine details [cite: 30].
*   **Guidance Scale (CFG):** Classifier-Free Guidance (CFG) is the hyperparameter that controls how aggressively the image generation process must adhere to the text prompt [cite: 27, 30]. Geometrically, the guidance scale acts as a manifold-contracting force within the latent space [cite: 31]. A low CFG scale gives the model immense creative freedom, allowing it to rely heavily on its unconditioned training data, resulting in highly diverse, artistic, but potentially inaccurate images [cite: 27, 30]. Conversely, a high CFG scale forces the math to adhere strictly to the target data manifold dictated by the text embedding [cite: 27, 31]. If the CFG scale is pushed too high, the mathematical vectors overcorrect, leading to severe visual degradation, unnatural contrast, and a loss of image diversity [cite: 27, 31].

## Next-Generation Architectures: Transformers and Flow Matching

For the first several years of the generative AI boom, the neural network performing the reverse denoising loop was almost exclusively a U-Net architecture [cite: 24, 25, 32]. U-Nets are incredibly effective at image-to-image tasks, utilizing convolutional layers to downsample and upsample spatial data [cite: 24]. This was the backbone of models like Stable Diffusion 1.5 and Midjourney v5. However, as developers pushed for higher resolutions, deeper prompt adherence, and better typography, the limitations of the U-Net became apparent [cite: 33].

In 2024, the landscape shifted dramatically. The latest state-of-the-art models—including Stable Diffusion 3 (SD3) and FLUX.1—abandoned the U-Net entirely in favor of Diffusion Transformers (DiTs) [cite: 4, 17, 34, 35].

### Multimodal Diffusion Transformers (MMDiT)

The core innovation in SD3 and FLUX.1 is the Multimodal Diffusion Transformer (MMDiT) [cite: 13, 34]. Because text embeddings and image vectors are conceptually entirely different modalities, traditional models struggled to fuse them efficiently [cite: 34]. The MMDiT architecture solves this by utilizing two separate, parallel sets of weights—one track for the image data and one track for the language representations [cite: 4, 13, 34].

Instead of processing them in isolation, these parallel tracks interact bidirectionally at every single transformer block [cite: 13]. The image tokens attend to the text tokens, and the text tokens attend to the image tokens, allowing the two modalities to evolve together in a shared space [cite: 13, 17, 34]. This deep, continuous integration is the primary reason why newer models possess a staggering understanding of spatial logic and typography, seamlessly embedding complex textual phrases onto signs and clothing [cite: 13, 34].

Scaling these massive transformer models introduced severe engineering hurdles. Stability AI researchers noted a phenomenon termed "attention entropy collapse," where the numerical values inside the attention mechanism would grow infinitely large during training, resulting in system crashes [cite: 13]. To stabilize the training of these high-resolution models, engineers implemented Query-Key (QK) Normalization, applying a scaling function to the vectors before attention calculations to act as a mathematical governor on the system [cite: 13, 36].

### Rectified Flow Matching

Alongside the shift to transformers, modern models have updated the underlying physics of the diffusion process itself, moving from traditional stochastic diffusion to Rectified Flow Matching [cite: 4, 17, 33, 37]. 

Traditional diffusion formulates the transition from noise to data as a curved, chaotic trajectory [cite: 33, 34]. Navigating this curved mathematical path during inference requires the model to take many tiny, incremental steps to avoid drifting off course [cite: 33]. Rectified flow matching fundamentally alters this by connecting the noise distribution and the data distribution on a direct, linear trajectory [cite: 4, 17, 33, 34]. 

Because the inference path is a straight line, the model can sample the path in significantly fewer steps without losing fidelity, drastically improving computational efficiency and generation speed [cite: 33, 34]. Furthermore, researchers introduced specialized trajectory sampling schedules (such as logit-normal sampling) during training. Instead of training the model evenly across the entire straight line, they bias the training to focus heavily on the middle portions of the trajectory [cite: 4, 13, 34]. The extreme ends—pure noise and near-perfect images—are computationally easy to predict [cite: 13]. The middle phase, where an abstract blob must be perceptually resolved into a detailed face or object, represents the most challenging prediction task [cite: 13, 34]. By heavily weighting this middle phase during training, the models learn to resolve fine details much faster [cite: 4, 34].

## Step 4: Decoding Back to Reality

Once the MMDiT has completed its flow-matching denoising loop and fully sculpted the target concept out of the static, the image generation process is still not complete. The resulting data is still a dense, lower-dimensional latent vector [cite: 6, 18]. It is entirely unviewable by human eyes.

In the final step of the pipeline, this refined latent vector is passed back through the VAE's Decoder network [cite: 6, 18, 23]. The decoder mathematically decompresses the semantic features, mapping them back into the high-dimensional pixel space [cite: 6, 18, 21]. It assigns specific RGB color values to hundreds of thousands of individual pixels, finally rendering the high-resolution image that appears on your screen [cite: 6, 18].

## The Anatomy Problem: Why AI Struggles with Human Hands

Despite crossing the threshold into photorealism, generative AI continues to exhibit a notorious, glaring flaw: it is historically terrible at rendering human hands [cite: 38, 39, 40]. The internet is replete with AI-generated portraits featuring individuals with six fingers, merged knuckles, and impossible biomechanical poses [cite: 38, 41]. This is not a random software glitch; it is a direct manifestation of how neural networks process information.

The human hand is anatomically intricate. It contains 27 distinct bones, an array of tendons, and a massive number of joints packed into a small area, granting it numerous degrees of freedom [cite: 38, 40, 42]. Hands can fold, point, grasp, and articulate in millions of varying configurations [cite: 38, 42]. 

However, diffusion models are fundamentally two-dimensional image generators [cite: 42]. They possess absolutely no inherent comprehension of 3D geometry, spatial hierarchies, or biomechanical constraints [cite: 38, 42]. They optimize for perceptual similarity based entirely on the statistical correlations they observed in 2D training data [cite: 38, 42].

When an AI analyzes thousands of photographs of hands scraped from the internet, it rarely sees a flat, perfectly spread palm [cite: 39]. It sees hands holding cups, resting on hips, or waving, where fingers are frequently obscured or overlapping [cite: 39]. The AI algorithm successfully recognizes the overarching pattern—that a hand is a cluster of elongated, skin-toned shapes attached to an arm [cite: 42]. Yet, because it is a "black box" optimization system, it does not explicitly understand the biological rule that a human hand must possess exactly four fingers and one opposable thumb [cite: 38, 42]. It merely mimics the visual texture it has seen, resulting in structural hallucinations when forced to generate a pose that bridges gaps in its training data [cite: 38, 42]. 

While recent models like FLUX.1 have drastically improved hand rendering through massive parameter scaling and deeper text-to-image alignment, and developers have introduced post-processing solutions like HandRefiner (which uses 3D mesh reconstruction conditioning), the foundational limitation remains: the AI is painting a 2D surface based on probabilities, not simulating a 3D skeleton [cite: 38, 43].

## Identifying AI Hallucinations and Artifacts

As AI models become increasingly sophisticated, distinguishing a synthetic image from an authentic photograph at a glance is nearly impossible [cite: 44, 45]. However, because diffusion models operate via specific mathematical constraints, they leave behind characteristic visual signatures—known as artifacts—when the generation process breaks down [cite: 46]. 

By understanding the root causes of these artifacts, users can both identify deepfakes and troubleshoot their own generation pipelines [cite: 46].

*   **Color Banding and Posterization:** When the CFG scale is pushed too high (typically exceeding 10 or 12), the model's vectors overcorrect, pushing pixel values to extreme contrasts [cite: 46]. This manifests as "color banding," where smooth tonal transitions—like the gradient of a blue sky or the shading on a cheek—break down into harsh, visible steps of color that resemble a low-quality, compressed JPEG [cite: 46].
*   **Duplication and Tiling:** Every diffusion model is trained at an optimal native resolution (e.g., 512x512 or 1024x1024) [cite: 18, 46]. If a user requests an image generation size significantly larger than the model's native training dimensions without utilizing dedicated upscaling tools, the model often hallucinates by tiling the subject [cite: 46]. This results in grotesque outputs featuring two heads, duplicated bodies, or repeating background structures [cite: 46].
*   **Background Gibberish and Nonsense Text:** While modern models utilizing T5 encoders excel at rendering the primary text requested in a prompt, background text often falls victim to the "bag of words" limitation [cite: 47]. Distant street signs, spines of books on a shelf, or small logos on clothing frequently degrade into alien symbols, garbled letters, or repetitive character strings [cite: 46, 47]. 
*   **Functional and Physical Implausibilities:** AI generators do not understand the laws of physics [cite: 48]. They merely replicate the appearance of lighting and structure. Careful inspection of an AI image will often reveal physical impossibilities: shadows falling in contradictory directions from multiple nonexistent light sources, reflections in mirrors or bodies of water that do not match the subject, or objects seamlessly melting into one another (e.g., a coffee cup handle blending directly into the grain of a wooden table) [cite: 48].

## The Generative AI Landscape: FLUX, DALL-E 3, and Stable Diffusion 3.5

The current state of AI image generation is dominated by three flagship models: FLUX.1, DALL-E 3, and Stable Diffusion 3.5 [cite: 17, 49]. Each model represents a fundamentally different philosophical approach to the same mathematical problem, resulting in vastly different capabilities, aesthetic outputs, and user experiences [cite: 17, 49].

### FLUX.1 (Black Forest Labs)
Developed by Black Forest Labs—a company founded by the original researchers who created Stable Diffusion—FLUX.1 is widely considered the current industry benchmark for photorealism and prompt adherence [cite: 35, 37, 50, 51]. Scaled to a massive 12 billion parameters, FLUX utilizes a hybrid architecture of multimodal and parallel diffusion transformer blocks driven by rectified flow matching [cite: 37]. 

By employing both CLIP ViT-L and T5-XXL text encoders, FLUX excels at rendering highly detailed human anatomy, accurate material textures, and flawless typography within images [cite: 17, 37, 52]. It is available in three variants: *schnell* (a distilled, high-speed open-weight model), *dev* (an open-weight model for non-commercial use), and *pro* (a closed, state-of-the-art API model) [cite: 35, 50, 51]. Its primary limitation is its immense size, requiring significant GPU hardware for local deployment [cite: 43].

### DALL-E 3 (OpenAI)
OpenAI's DALL-E 3 takes a conversational, highly integrated approach [cite: 49, 53]. DALL-E 3 is unique because of its invisible integration with ChatGPT-4 [cite: 53, 54, 55]. During development, OpenAI researchers discovered that diffusion models perform exponentially better when trained on highly descriptive, paragraph-length captions [cite: 56, 57]. To achieve this, they trained a bespoke image captioner to relabel their entire dataset, utilizing a mixture of 95% highly descriptive synthetic captions and only 5% ground truth captions [cite: 55, 57]. 

Because the average user types very short prompts, OpenAI implemented an automatic "prompt rewriting" feature [cite: 54, 58]. When a user asks ChatGPT for an image, the LLM intercepts the prompt and silently expands it into a dense, highly descriptive paragraph before sending it to the diffusion model [cite: 54, 55]. While this guarantees stunning, reliable results and excellent typography for beginners, recent studies (such as those from UC Berkeley) have shown that this automatic LLM rewriting can strip away granular control from advanced users, forcefully adding complexity to scenes where minimalism was desired [cite: 59, 60].

### Stable Diffusion 3.5 (Stability AI)
Stability AI's SD3.5 represents an open-source, highly versatile middle ground [cite: 36, 49]. Utilizing the MMDiT architecture and flow matching, SD3.5 is available in Large (8B parameters) and Medium (2.5B parameters) variations, making it highly accessible for developers and artists running local hardware [cite: 36, 61]. While it lags slightly behind FLUX in pure photorealism and text rendering, it boasts an unrivaled community ecosystem [cite: 43, 49]. SD3.5 allows for deep customization, fine-tuning, and the integration of control nets, making it the preferred choice for power users who require absolute control over artistic stylization and generation workflows [cite: 49, 53].

| Feature | FLUX.1 (Black Forest Labs) | DALL-E 3 (OpenAI) | Stable Diffusion 3.5 (Stability AI) |
| :--- | :--- | :--- | :--- |
| **Architectural Framework** | 12B parameter Transformer (DiT) utilizing Rectified Flow Matching [cite: 35, 37]. | Proprietary LLM-based architecture, highly literal interpretation [cite: 62, 63]. | MMDiT architecture (8B / 2.5B variants) with Flow Matching [cite: 4, 61]. |
| **Prompt Engineering & Understanding** | Exceptional. Dual text encoders (CLIP + T5-XXL) allow for deep semantic comprehension of long prompts [cite: 14, 17]. | Conversational. Uses GPT-4 to invisibly upsample and rewrite user prompts for maximum detail [cite: 54, 55, 58]. | Strong baseline adherence, but requires more precise manual prompt engineering [cite: 36, 43]. |
| **Text Rendering & Typography** | Current industry leader. Reliably renders complex, multi-word text naturally inside images [cite: 17, 49, 52]. | Highly reliable, though it occasionally duplicates characters or warps longer phrases [cite: 49, 52, 62]. | Vastly improved over SDXL, but still struggles with text fidelity compared to FLUX [cite: 49, 64]. |
| **Aesthetic Identity** | Unrivaled photorealism, accurate anatomical detail, and natural cinematic lighting [cite: 17, 43, 49, 62]. | Highly polished, distinctly stylized "rendered" and illustrative aesthetics [cite: 49, 63, 65]. | Highly versatile; excels in customized artistic, anime, and heavily stylized outputs [cite: 43, 53]. |
| **Deployment & Ecosystem** | Open weights (Schnell/Dev) for local use, plus a closed Pro API. High VRAM hardware requirements [cite: 17, 35, 51]. | Closed system ecosystem. Exclusively accessed via ChatGPT Plus or Microsoft Copilot APIs [cite: 17, 53]. | Fully open weights supported by a massive community ecosystem for fine-tuning and LoRAs [cite: 49, 61]. |

*Table 2: A comparative analysis of the underlying architecture, capabilities, and ecosystems of the leading generative AI image models in 2026 [cite: 17, 49, 53, 61, 62].*

## The Generalization vs. Memorization Debate

As these models continue to scale, a pressing technical and legal question remains: Do diffusion models actually create novel images, or do they simply memorize their training data and regurgitate copyrighted material? 

Extensive empirical studies and theoretical analyses reveal that diffusion models are primarily engines of generalization, not memorization [cite: 66, 67, 68]. The theoretical goal of the denoising score-matching objective is to learn the gradient of the entire data distribution, allowing the model to sample novel points from the latent manifold [cite: 66, 68]. 

However, memorization *can* occur under specific conditions [cite: 66]. Research indicates that memorization is a function of the training dynamics, specifically relating to dataset size and noise scales [cite: 66, 69]. Models exhibit memorization behavior primarily when trained on small, duplicated datasets, and this memorization is typically restricted to the low-noise scales at the very end of the diffusion process (the phase responsible for rendering high-frequency details) [cite: 66, 69, 70]. 

Theoretical analysis identifies two distinct timescales during model training: an early time ($\tau_{gen}$) where the model learns the broad statistical features necessary to generate high-quality, generalized samples, and a much later time ($\tau_{mem}$) where the model overfits and begins to memorize specific training points [cite: 67]. Because $\tau_{mem}$ increases linearly with the size of the dataset, training massive models on billions of images creates a massive window of implicit dynamical regularization [cite: 67, 68]. By employing early stopping techniques and training aggressively on high-noise scales, developers successfully prevent models from crossing the threshold into memorization, ensuring that the outputs generated by users are genuinely novel mathematical creations rather than collaged replicas [cite: 67, 68, 69].

## Bottom line

AI image generation is not a sophisticated search engine, but a complex thermodynamic-inspired mathematical pipeline that sculpts order out of chaos. By translating human language into numerical embeddings via dual encoders like CLIP and T5-XXL, the AI utilizes a reverse diffusion process to iteratively subtract static noise within a highly compressed latent space, ultimately decoding a brand-new arrangement of pixels. While the technology still wrestles with 3D geometry—evidenced by lingering struggles with human anatomy and physical impossibilities—the architectural shift toward multimodal transformers and rectified flow matching has successfully ushered in an era of near-perfect photorealism and precise typography.

## Sources

1. [Genius Forges: Midjourney v6 vs DALL-E 3](https://geniusforges.com/playbooks/midjourney-v6-vs-dalle-3-photorealism-showdown)
2. [Prompts Architect: DALL-E 3 vs Midjourney V6](https://www.promptsarchitect.com/blog/blog-post-2.html)
3. [KDnuggets: Diffusion Models Demystified](https://www.kdnuggets.com/diffusion-models-demystified-understanding-the-tech-behind-dall-e-and-midjourney)
4. [Propeller Media Works: AI Image Generators Comparison](https://www.propellermediaworks.com/blog/ai-image-generators-comparison-midjourney-6-dall-e-3/)
5. [YouTube: Samson Vowles AI Comparison](https://www.youtube.com/watch?v=AXv5sgIoPnc)
6. [Medium: Diffusion Models Demystified](https://medium.com/@mail_99211/diffusion-models-demystified-a-beginners-guide-to-ai-image-generation-2a3b7053d8d4)
7. [GoPubby: Diffusion Explained](https://ai.gopubby.com/diffusion-explained-how-ai-image-generators-work-fa4493aa8c0e)
8. [Dev.to: How Do Diffusion Models Work](https://dev.to/umeshtharukaofficial/how-do-diffusion-models-work-an-introduction-to-generative-image-processing-1mkh)
9. [GeeksforGeeks: What are Diffusion Models](https://www.geeksforgeeks.org/artificial-intelligence/what-are-diffusion-models/)
10. [Medium: Understanding Image Generation with Diffusion](https://medium.com/@d3xvn/understanding-image-generation-with-diffusion-78eea7e7d6f8)
11. [Encord: Stable Diffusion 3 Text-to-Image Model](https://encord.com/blog/stable-diffusion-3-text-to-image-model/)
12. [TechTarget: Stability AI adopts new architecture](https://www.techtarget.com/searchenterpriseai/news/366571066/Stability-AI-adopts-new-architecture-in-Stable-Diffusion-3)
13. [Unite.ai: Stable Diffusion 3.5 Architectural Advances](https://www.unite.ai/stable-diffusion-3-5-architectural-advances-in-text-to-image-ai/)
14. [Superteams.ai: Technical Deep-Dive Into SD3](https://www.superteams.ai/blog/a-technical-deep-dive-into-stable-diffusion-3/)
15. [Stability AI: SD3 Research Paper Announcement](https://stability.ai/news-updates/stable-diffusion-3-research-paper)
16. [Infograins: How Text-to-Image AI Works](https://infograins.com/blog/how-text-to-image-ai-works-latent-space-clip-diffusion-explained/)
17. [MyScale: Understanding the Text Encoder CLIP Model](https://myscale.com/blog/understanding-crucial-role-text-encoder-clip-model/)
18. [Medium: Understanding OpenAI's CLIP Model](https://medium.com/@paluchasz/understanding-openais-clip-model-6b52bade3fa3)
19. [YouTube: Explain CLIP](https://www.youtube.com/watch?v=L-gv2knl_jA)
20. [xta0.me: GenAI Stable Diffusion CLIP](https://www.xta0.me/2025/01/12/GenAI-Stable-Diffusion-CLIP.html)
21. [Milvus: Latent Space in Latent Diffusion Models](https://milvus.io/ai-quick-reference/how-is-the-latent-space-defined-in-latent-diffusion-models)
22. [Keras: Random Walks with Stable Diffusion](https://keras.io/examples/generative/random_walks_with_stable_diffusion/)
23. [YouTube: Fahd Mirza on Latent Space](https://www.youtube.com/watch?v=F9DLrka0MB4)
24. [YouTube: DeepLizard Stable Diffusion Theory](https://www.youtube.com/watch?v=IxHXQpk5kkg)
25. [Reddit: Stable Diffusion Latent Space Explorer](https://www.reddit.com/r/learnmachinelearning/comments/12v4f6q/stable_diffusion_latent_space_explorer_beginner/)
26. [Reddit: Anti-AI Belief that Image Generators are Collage](https://www.reddit.com/r/aiwars/comments/1exu7fu/antiai_belief_that_image_generators_are_just/)
27. [Reddit: AI Image Generators Don't Create a Collage](https://www.reddit.com/r/The10thDentist/comments/zmaa3j/ai_image_generators_dont_create_a_collage_of/)
28. [Depositphotos: Misconceptions About AI Images](https://blog.depositphotos.com/misconceptions-about-ai-images.html)
29. [Bob Mueller Writer: Debunking Generative AI Images](https://www.bobmuellerwriter.com/debunking-generative-ai-images-what-you-need-to-know/)
30. [Medium: Conceding Every Photo as Fake](https://iamitcohen.medium.com/conceding-every-photo-as-fake-9bef432bd972)
31. [RunC: DALL-E 3 Overview](https://blog.runc.ai/dalle3/)
32. [Wikipedia: DALL-E](https://en.wikipedia.org/wiki/DALL-E)
33. [Skywork: DALL-E 3 In Depth](https://skywork.ai/skypage/en/DALL-E-3-In-Depth-(2025):-My-Hands-On-Review,-Benchmarks,-and-Practical-Guide/1976472460575436800)
34. [OpenAI: DALL-E 3 Paper PDF](https://cdn.openai.com/papers/dall-e-3.pdf)
35. [Shreyansh26: DALL-E 3 Image Recaptioner](https://shreyansh26.github.io/post/2024-02-18_dalle3_image_recaptioner/)
36. [Notion: CLIP vs T5 as a text encoder](https://sweet-hall-e72.notion.site/CLIP-vs-T5-as-a-text-encoder-for-diffusion-models-df76bf09cacb425797640da86131267f?pvs=4)
37. [Reddit: Difference in how CLIP and T5 associate text](https://www.reddit.com/r/explainlikeimfive/comments/1rmt0ht/eli5_whats_the_difference_in_how_clip_and_t5/)
38. [Medium: FLUX.1 Encoders and Token Limitations](https://medium.com/@lbq999/flux-1-dev-encoders-and-token-limitations-8631c179eaad)
39. [GitHub: T5 instead of CLIP](https://github.com/CompVis/stable-diffusion/issues/35)
40. [CVPR: Scaling Down Text Encoders](https://openaccess.thecvf.com/content/CVPR2025/papers/Wang_Scaling_Down_Text_Encoders_of_Text-to-Image_Diffusion_Models_CVPR_2025_paper.pdf)
41. [Encord: SD3 Text-to-Image](https://encord.com/blog/stable-diffusion-3-text-to-image-model/)
42. [Stability AI: SD3 Research Paper](https://stability.ai/news-updates/stable-diffusion-3-research-paper)
43. [YouTube: SD3 Research Paper Breakdown](https://www.youtube.com/watch?v=6XatajQ-ll0)
44. [Medium: Stable Diffusion 3 Explained](https://medium.com/@pietrobolcato/stable-diffusion-3-explained-84fd085934cb)
45. [Reddit: Stable Diffusion 3 Research Paper](https://www.reddit.com/r/StableDiffusion/comments/1b6tvvt/stable_diffusion_3_research_paper/)
46. [ZSky AI: AI Image Artifacts Guide](https://zsky.ai/blog/ai-image-artifacts-guide)
47. [PCMag: Spot Fake Images](https://www.pcmag.com/explainers/human-or-ai-these-7-clues-will-help-you-spot-fake-images-immediately)
48. [SECOM: How to Spot AI-Generated Images](https://secom.plc.uk/blog/how-to-spot-ai-generated-images/)
49. [Leon Furze: Can you spot an AI generated image?](https://leonfurze.com/2026/01/19/can-you-spot-an-ai-generated-image/)
50. [Facia AI: Ultimate Guide to Detecting AI Images](https://facia.ai/blog/the-ultimate-guide-to-detecting-ai-generated-images-online-in-2026/)
51. [Medium: Technical Deep Dive Into SD3](https://medium.com/superteams-ai-blog/a-technical-deep-dive-into-stable-diffusion-3-f6e60e4b14e9)
52. [BentoML: SD3 Text Master Prone Problems](https://www.bentoml.com/blog/stable-diffusion-3-text-master-prone-problems)
53. [YouTube: SD3 Rectified Flow and MMDiT](https://www.youtube.com/watch?v=mgtd_EBxTlM)
54. [YouTube: Latent Space in Generative AI](https://www.youtube.com/watch?v=llSt93BoKkk)
55. [Medium: Latent Space Foundation](https://medium.com/@ding.zhongqiang/latent-space-the-foundation-of-generative-ai-models-9fa9a7cc4fba)
56. [The Academic: Generative Models and Latent Space](https://theacademic.com/generative-models-and-their-latent-space/)
57. [AI Prospects: LLMs and Beyond](https://aiprospects.substack.com/p/llms-and-beyond-all-roads-lead-to)
58. [Arxiv: Latent Code SVD in DMs](https://arxiv.org/html/2502.02225v1)
59. [Lensgo AI: Flux vs DALL-E vs Stable Diffusion](https://lensgo.ai/blog/flux-vs-dalle-vs-stable-diffusion)
60. [MimicPC: Flux vs SD3.5](https://www.mimicpc.com/learn/flux-vs-sd3-5-which-model-is-better)
61. [ZSky AI: Flux vs SDXL vs DALL-E](https://zsky.ai/blog/flux-vs-sdxl-vs-dalle)
62. [Modal: Best Text-to-Image Model Comparison](https://modal.com/blog/best-text-to-image-model-article)
63. [GetIMG: FLUX.1 vs DALL-E 3](https://getimg.ai/blog/flux-1-vs-dall-e-3-what-is-the-best-ai-text-to-image-model)
64. [OpenAI Cookbook: What is new with DALL-E 3](https://developers.openai.com/cookbook/articles/what_is_new_with_dalle_3)
65. [Reddit: DALLE-3 powered by GPT-4](https://www.reddit.com/r/DarthJarJar/comments/177e4j3/this_is_dalle3_powered_by_gpt4_prompt_engineering/)
66. [OpenAI Community: API Image Generation changes prompt](https://community.openai.com/t/api-image-generation-in-dall-e-3-changes-my-original-prompt-without-my-permission/476355)
67. [WhyTryAI: DALL-E 3 Better Captions](https://www.whytryai.com/p/dall-e-3-better-captions-research-paper-summary)
68. [The Decoder: ChatGPT prompt rewriting reduces DALL-E 3 performance](https://the-decoder.com/chatgpts-automatic-prompt-rewriting-reduces-dall-e-3s-performance-study-finds/)
69. [Arxiv: FLUX.1 Reverse Engineering](https://arxiv.org/html/2507.09595v1)
70. [Medium: How does FLUX work](https://medium.com/@drmarcosv/how-does-flux-work-the-new-image-generation-ai-that-rivals-midjourney-7f81f6f354da)
71. [DeepLearning.ai: Black Forest Labs Flux Outperforms Models](https://www.deeplearning.ai/the-batch/black-forest-labs-flux-1-outperforms-top-text-to-image-models)
72. [Wikipedia: Flux (text-to-image model)](https://en.wikipedia.org/wiki/Flux_(text-to-image_model))
73. [AI Free API: FLUX.1 API Guide](https://www.aifreeapi.com/en/posts/flux-1-api-comprehensive-guide)
74. [Medium: Why AI Struggles with Human Hands](https://ayoubaliabid.medium.com/why-ai-struggles-with-rendering-human-hands-ai-mimics-patterns-it-doesnt-understand-them-33d00e71c430)
75. [Avenga: Why Generative AI Fails at Hands](https://www.avenga.com/magazine/generative-ai-models-fail-at-creating-human-hands/)
76. [F1000Research: CLIP AI Hand Challenges](https://f1000research.com/articles/14-193/pdf)
77. [The AI Whisperer: AI Struggles with Hands](https://theaiwhisperer.de/ai-struggles-what-is-the-problem-with-your-hands/)
78. [Dev.to: The AI Hand Conundrum](https://dev.to/evanmarie/the-ai-hand-conundrum-why-generative-models-struggle-with-human-hands-21ib)
79. [Medium: Power of Stable Diffusion - Seed, Guidance](https://medium.com/@alimansour1/unleashing-the-power-of-stable-diffusion-how-to-use-seed-negative-prompts-guidance-scale-steps-7ec98f7ed921)
80. [Scale: Diffusion Models Guide](https://scale.com/guides/diffusion-models-guide)
81. [YouTube: Image Generation Parameters](https://www.youtube.com/watch?v=hrBIaDGPnK0)
82. [Reddit: Steps vs Guidance Scale vs Seeds](https://www.reddit.com/r/StableDiffusion/comments/wwv2qk/steps_vs_guidance_scale_vs_seeds_can_someone_fill/)
83. [OpenReview: Guidance Scale in Diffusion Models](https://openreview.net/forum?id=nfHimL6g8G)
84. [Arxiv: Memorization in Diffusion Models](https://arxiv.org/html/2502.21278v2)
85. [Arxiv: Memorization Behavior Study](https://arxiv.org/html/2310.02664v2)
86. [OpenReview: Transition from Generalization to Memorization](https://openreview.net/forum?id=BSZqpqgqM0)
87. [NeurIPS: On the Edge of Memorization](https://neurips.cc/virtual/2025/poster/115770)
88. [IMSI: Memorization and Regularization](https://www.imsi.institute/videos/memorization-and-regularization-in-generative-diffusion-models/)

**Sources:**
1. [reddit.com](https://vertexaisearch.cloud.google.com/grounding-api-redirect/AUZIYQFFUiZTr6PaG5yMKYDhxAnWcABjnJKH9AARNhfY81Y8FBvCPDW_8TudfW8SPybNK5-33P0W3UsZ_Yf6Lze4iAheyLKjrkEmg-8m6-mBT77GS79-GrXf-RH6ENMQUJOquIkF9kIB2mGq6dO8zEAwBxLxIlYeSDDdJfY45lnBf-L73a95GsP4vxFseuin70GgROcKGJIP)
2. [reddit.com](https://vertexaisearch.cloud.google.com/grounding-api-redirect/AUZIYQHELIpIJZxwQIRDTMvTjmhwry8QoCWvIUo-mSFopduNpdH4KJS7YpEf0lFC4cnk_IpGRa2MfZZMMSnLOL6P9Mb2XunE3ip3I57DmYDiX-DwCbyctixWUfJr3h69d9QkHnPhqVH7tlbTorQHhXb8TSdmMI9TWiUQHz71eD5iq-Jyqvk4HLuE_hI5WZzhWsoWMvIYMU7dZy087spMmg==)
3. [depositphotos.com](https://vertexaisearch.cloud.google.com/grounding-api-redirect/AUZIYQHiWWUbG4UQnAC9KTeft-1EViWHohfyAinZlg61ZaK-RarGYqtsPz7g2S8EX-qWXnn0Ixh0OiacE75l0TK3nQ8q8MuzsvD2ErnXC8_rp1Aa3fGpWCHkgJDp7KM9P9On_g78cjnxjxjdo50qyjmgQRbZyhM6niMXP7M=)
4. [encord.com](https://vertexaisearch.cloud.google.com/grounding-api-redirect/AUZIYQGnVD7XdCr2OegyJucBSbeJqkvijYpSPYFARvmuIe29NBB-bo9KjmpGY9Q0p6Hd-K7oObpbZmCVGWJE4dV_wlF4QQ2ayjwnxe_99qlPN8qQ-nwpK4PP2kRlz2gCwmQP3abKo4Bb2dgPLZQXDaaJG4t_d1gompk=)
5. [theacademic.com](https://vertexaisearch.cloud.google.com/grounding-api-redirect/AUZIYQH-ztfMoVcP59cyryOQiuHWRrpc-32Ld0AlUeZU8KJYSuLJSLOIx7GUBH9CnR1BFQvxlGllPBGyNjm1G-0xynW0Y4ntKYQ3oPmOFNYukjjBULAXBLf5bR6JeCXbUcoV6CyJTXrhh5JfQI_XgPNf4eEctCn5Q2LMmg==)
6. [infograins.com](https://vertexaisearch.cloud.google.com/grounding-api-redirect/AUZIYQGsqqIbug0C50Fhzx2_V8o9W80P-zby1xbInPDK84uvRTIav9dXdNlYMm-ARUTWE0AAzhkVNKSwUGh9PgvpBIt6hMLITUQFTAKkBJp-awVgnY55M7lCQhrtyol-gXUX76BZ59BO4BD74LzeQesRT-fIPRteSbch8Ab8PUZeTb-EhNMZP41IK9y_csCaLaBxiBbABkk=)
7. [kdnuggets.com](https://vertexaisearch.cloud.google.com/grounding-api-redirect/AUZIYQH2MOjCgv_D8gsT_1Nxx3DeOcG-0TcScZWQhDQA5NxrH6d9Z9otenyWRLMWgOHG8IZVO7lhfufTT4KtSZDhS__3i3VMDy0BDXbkW6ZDrkARiLbf8EXs5ar-BXWm5pqStvA36dpBrkFCoe-1_rlECQtmo8_2sjxBAxBQLYQFdNipXhR4IAePixRl3EcPiN21oUeZBoLxKj2zKRYwWHIP6MnI)
8. [youtube.com](https://vertexaisearch.cloud.google.com/grounding-api-redirect/AUZIYQHKQl12teb9P1iu670a7rVui_jgidoDA10kqfsVlvexE-7SppIwS8kQrJB1lg93N3euqlK33ETOWzZz6o2OL_WZS8CeME9ThwqUfjP9rGL6sEGM8g8M8q5fy30cRiTzcR68)
9. [medium.com](https://vertexaisearch.cloud.google.com/grounding-api-redirect/AUZIYQFjKMQh38pFt_IfA7EN00gomJ67NDJ-z9OBn2GvLRwQ5VNvsIhF5ocmx4VsHH3Xr8oOzpQ5Bg8hB7eblVunYi11MAJHSXqGkzJumhq0fz9MsMrHoQUlClgX6oWRdYfJmYw4aV6FsrcQCxm7KY0Id2wHYXAZvg55OAvAf7gfq5tfDGo=)
10. [myscale.com](https://vertexaisearch.cloud.google.com/grounding-api-redirect/AUZIYQGFyduz2CDXqOXaHZ9QNoTSJgP1u1K0wMD6ScMIN1rQjK0cm1byeBiGL2nVz5i7EnsOxZ5PAriHIHz1-rfFwnY00wxphbUnNn40TPhRU86_Pss5Sx4aAZl3UBLV9alzgXWHSNaMIsQWufXQVVyq9IDfIttqACE3AxdfQ-Svb_i1dKoV)
11. [xta0.me](https://vertexaisearch.cloud.google.com/grounding-api-redirect/AUZIYQE_wpeiINVhf7cuDagXC6pS8DhYlrTiJZDZnlLUu3Meoi0Y7d8bjY2lU3Pw08MEPQdFK9l1UHH-duaTkn7tlkS6b684MSbdqybaqvIKbjcI9E9mOG8pChToCxLqid0G5Ahc4UVRzvmot6utAfAt7mwa_INTJww=)
12. [notion.site](https://vertexaisearch.cloud.google.com/grounding-api-redirect/AUZIYQELdf3ck-uWZhOVm6LOysCqBtZt05VGiqiJXFNCbE-H22CR8FdIybILKu9SDAFNvBPfQE7we85nnLwfWQHpgZgyUetiwHz5CzgsOKfWIPmof1IdDfgTDv6a7WH1kCPnMlBKclhentl35aK8jy2111m4MAmtZipbYTFx14by3nAJuWaFDBtTCb13QOQesG5w-R1IpVUrKfNtjk_6TMC_mHNyHWIblTcPx2En9OWyHlN7PgA=)
13. [youtube.com](https://vertexaisearch.cloud.google.com/grounding-api-redirect/AUZIYQFtntfU6KX3tz8sy1gLQsLRJkyAHiA3N1K1BL-sxicMPnSP-Z35elbYmRjQoJRrccn3f0e1NNna1e04QjaBF8Bns3uAd-4R3cIBLeeRXpPyhS1tuiwiA4yzQc--16LYNIH3)
14. [medium.com](https://vertexaisearch.cloud.google.com/grounding-api-redirect/AUZIYQFhEg_ZYALOIdeqXxsV2505QcQZGymIVVbbe3j3UdoLodKoprAxk-mN58-fwG-IHj3I2oAcbBpDhC8S37nD6ehDiqoRDlSHQijhVPz1hrrr0OwZCccf6qDCQdfN4mKfEhdf0_naCjcVDOs-Wp4EqRsER-PmMRa9Tqw1G8mX27TkH1cV4cXTRrA=)
15. [thecvf.com](https://vertexaisearch.cloud.google.com/grounding-api-redirect/AUZIYQEw4ND9v5RQAIiKy0OYTKoPgSX7rVJq_ohZ2m_oGrzBEY8eAA7s8XApK6TKknAA7HaoQ09xVLl3QN7hNNl98tFN6Hhdn9Zgm6JjXYbNy81-iC_ULvtOPo7fzg8J6H7N7n8GQd7mgRqNwRwZTWtTY9U6aVC2vzmoHoz-VoIwjb-BnBArk_Zq49s2YCRAMR1jrgUseJK4xHq3AxFROm86ieZ_q6IFvRJKZKXs0DfnqfHbQhhYQIoPvCYVx2m3qPwO57o9RQ==)
16. [reddit.com](https://vertexaisearch.cloud.google.com/grounding-api-redirect/AUZIYQHVkUfVH4K-CvgHdcK1vyJ53h_x2otOq_7sGzZXDPdBPyIYUUXO3CExfA5CqykYNrSH9Rv6GggLkFtbN3XQUTeeS8t7nCTwpKtViglkiyx1QY4FW9jt5gAmCkVpP8iFiHAU1wYpJqQvBCZZvIvPjLjBOMkUHyB0Z43TfK8VMgJiSHfvTQ_BL6OxP-mYyBxmoC2KOFS3_iqO2MxiHfNo1GY=)
17. [zsky.ai](https://vertexaisearch.cloud.google.com/grounding-api-redirect/AUZIYQGxfmREHGT5SxPNlFcT_gmPFHY4ogvsAM3oPDkmM9mzpemke-5wuoadh9mYMOC3CZx1mXAjnXtRfZxRiQ3x4RRrNwkOA9tJEeDqQtRmJXU0ObcQ5HTbxOs5DL4F1H8jbLg=)
18. [milvus.io](https://vertexaisearch.cloud.google.com/grounding-api-redirect/AUZIYQEWr_CFEyXAczQWQ25wBWNuYCLzmB6tX3zsItok9QeenW8LN2chDyPKwsN1EJPzVQRFy63_kAFYLlONB4ASiAsSShXN9NqPlM42N7Z83OzNP_R6PdZaDYlhrY4KrK1jDmfiG95PxmYhxakzGqtzy6mbNvCoKgcRcDNb7WQsOGf9JPXx4LgV5ZBnf6RDLEB0hs8EJJNQ4Q==)
19. [medium.com](https://vertexaisearch.cloud.google.com/grounding-api-redirect/AUZIYQGNQx386noorblYb-LarEpds0Zjsn0OZYkRJVjGg_tsVLkqLLum7n5soz4RNwfOJe0It0UdJDOmbh3TomST1WugpXlvFT7rij4t7KvbQZWItUEdO64wCmlabnZfsLxIkvNJrmxwSM2SuoZJrDk7ZC43_zz2Z2fnTP3_694azs9S3e-4VUhAxOVZTsglhzsjkyU6UnOZbuUrPXw0)
20. [gopubby.com](https://vertexaisearch.cloud.google.com/grounding-api-redirect/AUZIYQG_56f2bLFmSfuRigroHNB4u2D6uSEgnpOtDdY9HNepbW0_icu_IEcD5V0NRs6FXUmxFdVNosQRuuq5cZsPuaCQ4vS8vSdw1EszEQ0-60nnDeeYRwp009SnPuQk61DAlCfet0wVOM4mGHJ5O5iiGwWM1OpFu8_36UGZgLhykf7sDoZd5WUG8lj7mR0=)
21. [keras.io](https://vertexaisearch.cloud.google.com/grounding-api-redirect/AUZIYQERP_aYoUgsi41oR2Dl1KgFjXqKEx67iIIwnAoSyRpu2WHVlOOrC7A4GsX3FG9IxRwLmnX4OploUlGC6vtT0QLaKNVea7FtEHBJGygRK5OpQG9vbvFdFlPnL_qrOZRZZICWydOs0xiLAumzTsCYoKQk6y0vuc3Hzw2B0HGwVGQ=)
22. [youtube.com](https://vertexaisearch.cloud.google.com/grounding-api-redirect/AUZIYQHZdAQGQQGUHv3oC3txtu7vO6AL_bhSTCBooTIbMLRim24tJ3PmDyWaRykI-yJDAqYj50r2yaUeaVlzg_l5QNnMQkl0h9ROSDr417sXTJPrxZTWQ0bLFz_wdd4NK6QrQfRn)
23. [youtube.com](https://vertexaisearch.cloud.google.com/grounding-api-redirect/AUZIYQFAEVRcJhuY50-AUeUJZqZqGzfmhl9XOpAmYjVN1L72AG__gDtXWrWn7qLPBIk9Ys33uzgrW9X33Gq0fbavVJKEXdvz0wHH11mMORLwUjGSxicQ1pDF6Px9VAd7JpPITWu7)
24. [dev.to](https://vertexaisearch.cloud.google.com/grounding-api-redirect/AUZIYQGoaZCDAiEWkqLUYADSmNo1qCSuIlXrX5nBJ4JUku-kZT-BRw8twJ6BwBIy_1Z8-poRNW8B4EESMN69zmlOdeRUU5HOrPIIX9VNyIlSBzWGYnwgjJnLFRa187ysPu0S48p1WVWOMyw1iPo4WDDE79jRYPCSeM4GFk_LrVp5ikgGWdaydGR2s2zufZsSVX8anLuXMP0ZOdHJXBHSUP9wuqvHdTbo0c07SEwmWg==)
25. [geeksforgeeks.org](https://vertexaisearch.cloud.google.com/grounding-api-redirect/AUZIYQEXZrTwLO9EomoqEhgteyjsSwWosO1lnwZ9bMI9-1O-wBqVLlSnrGT4nX0D9_M-T-Sj7a0FMNWbxHPM6Dnf5xDGB2nlUP_pQ2lRWekYhA7kW2-zzEN8H2GaPuLwy6EPC_vP4bsakJMAABDWCUrx4hVDERG4yVu-mOBPUydtM0D7tKD5Iz2yww==)
26. [medium.com](https://vertexaisearch.cloud.google.com/grounding-api-redirect/AUZIYQFJB41XVCxtuXJBRglIpfFCSHn4jcSFe81wtsS9knJj-8dSr6VYcv7AFi9w3kzddJozBKtvfIyZGvSYBFV_t41O7bELgOhzjEkZHHotDZFJYJiWRR04QBWWSpJLnedu-sn-UlifJQzs8Ch1VkyseBrWT_WxJuqkdrQs0Df1b6AJRv66NP-z9_6nwRXyIAsBxxh8k0D_d86g-3QeMFeP_qnoDUzzkIyS0Q==)
27. [medium.com](https://vertexaisearch.cloud.google.com/grounding-api-redirect/AUZIYQGN9WqG3srAQUD7hvp2ztq0Tszwtu47rcd2gPRs0rwE5tvJFe8Fm0AQGZbJs1-0ZiaZBekAoafAlccsyrNwJVhCXSmDi8UDrcgrLAUKEyYTfbOH7xnRieK8gHKsH9qyAsYIQLCPxtXjdJHy3e9BoVDj8nhvd6K7JplF6a8a-uXcvDDfuIUf03I41P3xTJeVps7vFFG3FHSqj5zi2d5IsnN8ga7Qz6BJcl5ElZ0PzzwxTJUfOhaoRcp9bflBDAMXPpXSYQ==)
28. [scale.com](https://vertexaisearch.cloud.google.com/grounding-api-redirect/AUZIYQEs-8pHLi8SsNqHRvOMxGXxqxprSyjxmnJwTE9KGvlucrHKofKRgOCHxTX_Y6rYflS0RYONYDAb1tFJ1dHEBHHsY9Bn4mDz2byqaiFP5IEwjwyQ8YSuhyF7BDe4fGrcC64etbZqkg==)
29. [youtube.com](https://vertexaisearch.cloud.google.com/grounding-api-redirect/AUZIYQEXlZoldZJ-Qwy1cDo_hKCTFWAAZIixj8BMDGLocPSqZ2LYMLJnlvC50mpMUKjX2GyksLG711vS2AN_kSeFrxHd3_xBXZ4nvViiTVlSTDVdrwnQOgBG4tADGyLuHp7HIR3a)
30. [reddit.com](https://vertexaisearch.cloud.google.com/grounding-api-redirect/AUZIYQHCauoRhluHhWNj1IQdLdNDnhCQRSb9-HcCEEVxznuf5R1UCkrZf-X3R72v4bCZ4z5-pH2ULp93bZvrhYKaEn6sev_qAXiLgstGNPHYQ8XXqO0XII-5P3VoKkv5ygasCJTTden4K4trsPZj8SwWO6fhHcBqMOuq8za4RLK_tb07d1Zhc-bgc-c--ya1DrZsZjPMR9m076xCPurHmtPuYVXlfQ==)
31. [openreview.net](https://vertexaisearch.cloud.google.com/grounding-api-redirect/AUZIYQFZcCRwYalcwUP2CuU5vI9ePjfz3FOqLuy7toBwuIlIAfpewZpBOc8Wv8AGfcSvlhFf6y5hIV9AzpRQE6V434AwrZaW-t4I-9ihZ37YAn8sl9WBnYyHzt-azxGH5VuUw4Q=)
32. [youtube.com](https://vertexaisearch.cloud.google.com/grounding-api-redirect/AUZIYQGofkPLCwjQklm1guT3t9rD4rLKrERrkZkvlgbMa-OtaG2dr7g-w2YpF6QT0A3-uYFiGhYvBEpOVKMHJt0wW5VltgXV2VCVO9TvFafyziypAep0lZPbVeTGKqWOzq4_f6Rt)
33. [medium.com](https://vertexaisearch.cloud.google.com/grounding-api-redirect/AUZIYQFTzn0WZQmyOnB7U8q3OBD4y_O92-MNnCjQc8m4KLyjgigljN-QpLOH0EMoUEi6nnatkOn0eCJSZsrfM7PIERsElXY4oo01-r7kb-_zy9tnoN31SOTjW_FxTm-5XGd7XF8Eep1X13s0Kfsg_anQmy3JsKcC1efgnaq8mr88ZWEA_cg=)
34. [stability.ai](https://vertexaisearch.cloud.google.com/grounding-api-redirect/AUZIYQHA-ukt17CarpMYF0wsSBmXIFIE1g1okD4fcIloUSLiafd07aKmPO3_yMrvYD1lgnCqMC_KEaETZh4dUQksvZCCq3_s3d0wWX40kYhGDji9h0UF07xdYJ_N0V-W_L-MZiZ2r_TwJz0DuCBYI7RsapACZitGWdykT9U3)
35. [deeplearning.ai](https://vertexaisearch.cloud.google.com/grounding-api-redirect/AUZIYQHLmCszAX4hlOnMuimy9S5giTOqyd4MFQSVeljWLnIqp9PMlP57ZOLqf3DDJgODM5eYVLeAgZy-rSuxjPMnqzhFfPZRh9HThFiRYyJQsuLE968wZ9UdHTQOONFC_hQZoaDUj6QIZMiKgNxXv-x7AeTs_5outuqIlg99TO5zinQWsG37FDaDR3DKN3Y-2fh0GE_zaZk6bPWCYUg=)
36. [unite.ai](https://vertexaisearch.cloud.google.com/grounding-api-redirect/AUZIYQFht3SrPDAF3JKNfEDBruaAZTWFrGGB2MrBgcBNNj31DY9Mfh_Ag5K_5cG-QAA4GN8Icn7XSo1k5wyJnVq_RX4irqNzl9v-MinR3ZrorcgNcXHvQhaaoTIbXW_vGcku8tMnXymHqJ_S2F96Y1MlZOOn8Wh642LStOEOTn6TPVsb3umsI-7QlR0ocKKO)
37. [medium.com](https://vertexaisearch.cloud.google.com/grounding-api-redirect/AUZIYQHZK37wtJIguBVC42cY1iQqYEL-oeacLPKAKRlbZnmAo7eokRXB8-y8fhKLM6dLlZzRMCxrqPiD7V-8Hnee6PQG2mVLElHe3cYKpix9fgGLDfX_VVsFYTVTiD_OPiAkqCl60FP6ttzZ9UPm7je9xyWeGXx2hW3OOT4wpNW9-abgj6mssgfCvWkdpS3cDIdGP7pjR3EE9hW6MFU0OCsZg983pYZe2q6Y)
38. [medium.com](https://vertexaisearch.cloud.google.com/grounding-api-redirect/AUZIYQEgtuZGPRR7KdZw74kq74MK_dEtPAO8B30jdHu8hpjNVQMAIkBj1wApCPFkVN37-BLYoMNVdknKFVS95OJXznpaigv53wcFMJC2NKVB5du6Ex6y7oIl3ag4Kmre6tmR-mEErMfhLhW1oe0t9YVj-q-FWp3SHsgrV7onp4VH4VF4TnYpo-GgjPXRvEuHqSD-ogcYTQr5JbXpCHtjCIyNBJisVaoHDA_90ufRBCo5NqZKyTsSZcl2e8yiQ_KKRw==)
39. [avenga.com](https://vertexaisearch.cloud.google.com/grounding-api-redirect/AUZIYQFL22PQPW_hCM14XvWE8IpphG1yr_BdpZ2vgzIbqRoNnrvlBqzHaPd9f2CQCP_kIRAdVbyWF-IyhmMx1v4k3cguZjOvedd4pojRslpKrFhvsGPr-eGOWpO3mdipnNAVCLKLbxbxbURl_NwrQEtFaGtk6C6G3ciLr_YHjm6GDXKQXVPYTxIDZeDr)
40. [theaiwhisperer.de](https://vertexaisearch.cloud.google.com/grounding-api-redirect/AUZIYQEZBVo73q4w-aw8bca4vIk0kHKxiF6RiIFpYAgET4vbcowZmsZrlLCfy8WfERarE-oDVZ6-u2oaQci-3Ne8UKWn7IzmXdNJ_A6WdtyjCOmCDqSm8ebYx34DmhXGSbGsvgh4gTLyIzxcYGwVXimh6EqyppEx9o73OKlDNjzzje-OSmU=)
41. [f1000research.com](https://vertexaisearch.cloud.google.com/grounding-api-redirect/AUZIYQHIPJgbXa3X1_nm27vc0Q8XG2hG_zjqFE9dZ94WaPQfhjyHR8wK-96cITswhDgunjl6KOephHxKQS-ayTK6srZQzP4uutQIqLOy-3XvOPEN5gnSQeEHdVq51xUZ8Ft3Gv5IuMM=)
42. [dev.to](https://vertexaisearch.cloud.google.com/grounding-api-redirect/AUZIYQG378S4peyvN95cpCO-6yBeF7OsJ3y6YWjyVmFi_71fQdcAcicDqy6wcXO14gmEPFf3N2kG9LY_8fBpF5TcpoaOA_bXrta7YgD18MYjwqt48dWoiLKfdPyCbtp-NLM1CHQiMdXtg2jAxrpP0gnXSkmTfPigHIRAE0P3Bd7pDJk3oLMgX_I75Fooe4kkJPnah0DXI1rDjHGxqX8=)
43. [mimicpc.com](https://vertexaisearch.cloud.google.com/grounding-api-redirect/AUZIYQHqJcLOzAc5l_bjPNIvQstMPG48R00dzd_US1zg3NxwDBOzbobvIgcYOvp5QVyXpgsBUt3kAKs81eDRMUw0PFM8qLXFWjiuPlX1FBDP9eROyJJaXZp6FFK_HVom95Po4lcOqVlxr8Q0Ll5gNqPZoNb61T-RWZYQQQ==)
44. [leonfurze.com](https://vertexaisearch.cloud.google.com/grounding-api-redirect/AUZIYQHvPa_56UGAypR35oCzD7wK0V6_VHw-2W6XfK09iS9jDwgFqRigMRD0Xaj7THmOfDH5zIAHnR_x0e0tQiEmOOVv-6K2m97iM6jjuhxyKb9QaADBwxqRGgB_OFON12eCuhMBkLMsPCBCp9ia4IGKN_fgv7fUx9rYFCT7Og==)
45. [facia.ai](https://vertexaisearch.cloud.google.com/grounding-api-redirect/AUZIYQFwF6isJUYjVhIcEwwTxCBLi6zU3JHxD3xhAbTUTZeAug1UIecgffI4JWqd30D1fJQinhXq3Htmg5c2nj0iVkIYg6CvWLOQuQITaxcO_rfh0e8elHIrUQwIasYZ6Y6LpqahnR8daZbCQdOpSRxkA9BL73kdK3g8i7z6pKnxLFutUIR3clmPFPOLfB7cxz9wjQ==)
46. [zsky.ai](https://vertexaisearch.cloud.google.com/grounding-api-redirect/AUZIYQHAp-xfkLS3XJqy9pMA7_pJxPE4l2Lo532Wjdkm4znHW45zNKrXjvfrU1bSOOo9dYygkYeDl02oAJQ2JnmfuWxUHIbwOk2YtrUGAl9oQptYnZETfsLKK503Stniw_sDMUbFRts=)
47. [pcmag.com](https://vertexaisearch.cloud.google.com/grounding-api-redirect/AUZIYQFezAKgH9oaKT7XprD3MvOTrgH095orrnUjELCmJ_QqdLNJAjQ0O2IK6v4AQZ7CAJ-r3ZzcwxE6JQ_v6_nuIMy0GYmOBNZoPIHo5aWPO0LBibkyk3TF519hXUVGVE_8iNO_zNCDcyvNvxX04x7-_HXuwlxZw-R0t4C3OhtnbgvRrulCC2mb4OHtJ1rYefUZ-FFzxNpgmntsi7T9CQ==)
48. [secom.plc.uk](https://vertexaisearch.cloud.google.com/grounding-api-redirect/AUZIYQEkyA3h83i0wfwV540goPOPWBLhTomnvgw_fBtovyTscEBZnLyWc4J7snu8GyiS6-eIsulXhXP01HHAGg_JHgH8t79LpdsZZ5qNtbdixC6PavwKRoah62JGRtRoAoY70Sgd8H6sJZN6_mNJ11fSnjvZ)
49. [lensgo.ai](https://vertexaisearch.cloud.google.com/grounding-api-redirect/AUZIYQFRnHKvW8XO0rdjQLgtwcR4SkF5HvsoFBgS_8zlyWFlntYySaf3XOvPaoyAGBeFJZ1XFpZ5fvzNE9pcliljWCnucHg-PfsDF_KB8GwDzf5ogSzcKf9E3xs1zBPoH39i-QXjwMgp1HwedPZJ-HtuqQ==)
50. [arxiv.org](https://vertexaisearch.cloud.google.com/grounding-api-redirect/AUZIYQHt7MCSC7qcAL5a-pvRbcaNlXTqNend9MNdWEDN-QFewq9vq1xVUq6qsrtaxvpG6WJU52-_5RnIQHAyAHjmePt1v5bhd9Ke5Llqn6JjhJY7atuh4nh0UX_eFQ==)
51. [wikipedia.org](https://vertexaisearch.cloud.google.com/grounding-api-redirect/AUZIYQG3nlGLUHUe_hb4aDJZQ1hymvuixmfx7ydtTiibKXuKR_ld4t3AM2GrSS3A1oiZ1YdAvuOl9IauTav9IQm2aQcv_dPcBmFPM7nfe-4bN8FxChKkukSyp4E37TOIFN1ktUPMSEhx6P1upXPvaEad-Q==)
52. [getimg.ai](https://vertexaisearch.cloud.google.com/grounding-api-redirect/AUZIYQFwoebW31eol_ojrYyEJhlglFCqWi2rwdGQspelRBZnsiPuBwnPkt4w8qKr6TMuunQ8ajx5vQjM0UaL5XDSDk4u2PB_QdhUUcHohXvKjcdliUrxZeLETmJYE1ZhcR3Q5E6WGTb0LFziaqmh--SSciIQHcYT5r-6JtsMlUKP1Utjo4rvpSRV3Jo=)
53. [promptsarchitect.com](https://vertexaisearch.cloud.google.com/grounding-api-redirect/AUZIYQE9QKPUbeztLWwJSNapZK4Z8vsRiSIdlDH2OgTj36FXspg_ury0q3gqNJjbtDK5gutfp5WVY1f7e_jJESDp9B6cyVKOEOU_xtvd4APycmkm6ZaUznAArI_BOxEMQKGsuedxsiGaZUTp695bWBk=)
54. [skywork.ai](https://vertexaisearch.cloud.google.com/grounding-api-redirect/AUZIYQEOk1DY0ZHwyChzScDuDyicCQkDcDKAYFm7kCYuB7vUjftZQnsC0N92NTZFHEYmkjXcsrxcYCYoSDFFkZpeom1e9OkukzfF7J9fBHasusnCRYM2bvvw5R4f0dKptkmo2cXrsA4fn47c0hyIsC5vq8yMYmMMaIbbrBcH6AJAKJWQxfBJoiAhfxp0zrRezOblcPfcA0mq4fI52Zs8swz6B4cW5AUIo4mDltUQ2imbav5S-oZf0xD3)
55. [whytryai.com](https://vertexaisearch.cloud.google.com/grounding-api-redirect/AUZIYQGpqHPQsHgbb0rQ-t3WOQTsQzNelDObdG-kXJWHJMI532glO_hfXfm8cesnv2CeWw_9yy4BpJtNKYrPsdBywUGAlFsK-5zTaXSEmP_CQRa1Euf4z1FTTw5rOZUh96AzBsPFO8FuGXOq7rCZ6QnGc3rB5CjvTR3JIKSRbzwJzyJHfA==)
56. [openai.com](https://vertexaisearch.cloud.google.com/grounding-api-redirect/AUZIYQGpyisxJy9fOdIQghasVQMKPFVJHY5UZybGoztrun3wNUiw6gYVxIrDUCPq6YPKYmWxv8GYDR1qu7oLBfKV16999tPiAUiAuF5TwDztXlX9LFwcA_bMhJXlUxfA5mANbmc=)
57. [github.io](https://vertexaisearch.cloud.google.com/grounding-api-redirect/AUZIYQEntWQC0Qwn6EMXJgHtVr5aItE6FcIAW_M5NbM8TgxWHKDsigoSVTW2UGSciS23zsCVqNYJ5FnGTP2Ath0zNBdlBFUc8E5xrSAjJqgxqHT5lQjJdZRBSES62jUc87HOscPEZ8RtcgTDlYPCr3zNkmCXwm3dOhzPrDSzKRrWvg==)
58. [openai.com](https://vertexaisearch.cloud.google.com/grounding-api-redirect/AUZIYQGEv31tQ9MbwxxjI4KP8Nz4FLYqjuLHMDxp97O2lUEWiVk1oimFj8CSq9OocC6WBTa3dPCVjRSmhSjA4iYdjtsxZs7WtvttwH0LMsqjUUF9GeB9lbq4M_E29DnWpX-SVpJIapTL8VMngf72q8n9bqEZXEOZAlH0iSrrAtkOfQs=)
59. [openai.com](https://vertexaisearch.cloud.google.com/grounding-api-redirect/AUZIYQEbKtB7wwGBOOHikFXWPzqTjzu3K3k_BKOg6fDh4tu0tdvPG_QRZY8JKgmlAN49u7MJcOdmcH-OgBIc8apzgTqsKnjezrSoBbqZcIDDmMj8yhXvm0oNZmugIOAZRjcjCb3yzuZKJuqgkb0KFuozUOd_9HBpqxlX-FvHJm511AfnDN7dsS-_vXcdKYXOtHMJtF9Cfd5wTQz7sS0OuH0RXQFSC4SwXYeOsbhUsQ6fKg==)
60. [the-decoder.com](https://vertexaisearch.cloud.google.com/grounding-api-redirect/AUZIYQEFsLoAoJXP4-1tfFGOQe00BV5yw0iMLRUbK44mYV7qIzaTS9vguldtXqM4XC1aweYhF8ugEu0MlKzUreSRTHI7y5m-jeMJlJGAVUDyDFoUykbui5xYvV4i9o9hWeslTneuo03rhkAuJJnopBUS7lha5Y9HKPLm-Na0ov8GGm4MxgRy-6EnZqQXELCLOja0NdLoTtj2G-8Pcugm3xc=)
61. [modal.com](https://vertexaisearch.cloud.google.com/grounding-api-redirect/AUZIYQF40LNkKxtMRUlhO2Tyzjjiug4Xhr_gvbzpkkw-p7rSBjpCwMNw2L8XPEs2LvRSR2FMz6spm4wWerPC1Am1M_EayCb_zt7Gdq2e_t2sMuE9mVFzudT3XPcEcvQjllPAHf7ecjZa0sXq3CHiAHQA)
62. [geniusforges.com](https://vertexaisearch.cloud.google.com/grounding-api-redirect/AUZIYQG91geZSzSPRR5FCsXwVVsXlCJjQS3EKuiZkJldsNNTdWwz-fnOAuRhQbXYHdgH7ilDcRZJ_A9j9mqFuPo3LoWk0R6WPMeizE_WU9H3wwdfKhDqrJcS1giKTZFqxzrAzO34STlEh26X1vKfqSvVSnLsGiYXmPrrrHK2RqD-jDGOxPwgveie5A8=)
63. [runc.ai](https://vertexaisearch.cloud.google.com/grounding-api-redirect/AUZIYQHnZ9qeg_qP7WZeoi71RZswEdHWRRpg-qi_0tNxPwSm3v3HEn1GTGrPcFtcLIvszm02i6vZ0vEdQwisH4xa7195-VRsH2IGg8IuNn9RZ0VJ_ANg)
64. [bentoml.com](https://vertexaisearch.cloud.google.com/grounding-api-redirect/AUZIYQHnzKS8_7f0Gj5R04CJaS2mNCidb8vTkJYVdnIatD_Zw9Wbx9iwkuk2vPmve6vAEwyG1gVNjblvOge18n46BgIDi0iZDTN6M4Tr5jwWTCV1HWiK7xO42xGeQzrNkjmBJ7o1NoYKSARU4WJdZ06_3cB0wN5k1F4Wi20hS35JGTi6Pg==)
65. [propellermediaworks.com](https://vertexaisearch.cloud.google.com/grounding-api-redirect/AUZIYQGORx15mQGcrmdLxHcfGeZWUVpmf2WRMkBQvtQ56mKo80bn3UoG_msfGAGkIv4US2tBVJSFk_dCqxFtZRQbolgY--eCm1nVlY7mSDi2IdoQ03kB8OL_3phLGiOp0V-oNiG2f4Mh8Hs48orXK3mGXmAJvqbEGqcHmDhIgUwxwchSx-HZm-y-T-VCzyTsOlQGJpyxKrF9)
66. [arxiv.org](https://vertexaisearch.cloud.google.com/grounding-api-redirect/AUZIYQGUZLsoWmrDgps9HNz9FIeCUodWf5seXYMB8UEh9wHpf7hKo8galt3-WjpgXMH25ejEKspmY74wd__LT8_LlLdBBT_A9LCN22NoanHLkK2lab-hyWjilnMHDw==)
67. [openreview.net](https://vertexaisearch.cloud.google.com/grounding-api-redirect/AUZIYQGmZ7GQrdNv_fswmyLQ4K71k_zL7RWV4wL4qq71xyzSQ5epe2PY2zELccNg-s_FMcrhx8IQyZNQNBEAyO8sDImsyxvzbD7PWr6PcmRHzdS7KkNj8Tn7RIooOkCRX9E3Zns=)
68. [imsi.institute](https://vertexaisearch.cloud.google.com/grounding-api-redirect/AUZIYQFXk724WiRIgTqQ0vqqOYibdQC1EM6vP3CmHuzUB7ND-bxptqUyyBG9yKipGWaCvqPJRXR4Zavtdg2tMCd-y5uVkCB0XCmlk_51NvjUEbjTPm9kPD32i73cYWVyvNpHZ_k2V8XpzJfEaTOl1jeH707oD2q2o_DSirP9W_YfawQXLXRBQeVSEnjoe0CLGhEnSlyP9gINIfVT)
69. [arxiv.org](https://vertexaisearch.cloud.google.com/grounding-api-redirect/AUZIYQGhc2UuCJpsX2QkD5wz-YvJ6d_X80JCvmGNSLOz4GnFFmup3CKCRm2PNlFu1PB5nmDV-0lSXmIxFgmKhgdj1VHxlJh6RWjNXtHUZxraanx1LOG_1_EktnHBWw==)
70. [neurips.cc](https://vertexaisearch.cloud.google.com/grounding-api-redirect/AUZIYQGNvpwq0rLP57pTF27HuOKfCpGUzSSu3oqTqRgLrCeMw3LMGe-4BKhE4TJZWwcpBQsOlBmHb8iJRRw4QNrMFW_TuIobIPDsWm5Trs-Q1xlH7pYGD64TuMvC0qknX004PLRvWnw=)
