🐦 Twitter Post Details

Viewing enriched Twitter post

@llama_index

If you've ever worked in or around legal, you know that discovery is where document parsing really gets stress-tested. Low-resolution scans. Black and white images. Handwritten annotations. Charts buried in reports. Files that are technically PDFs but practically unreadable. And hundreds of thousands of them. Traditional OCR tools struggle with degraded scans, and anything visual (photographs, slide decks, tables) falls through the cracks entirely. That means your search index is noisy, your recall suffers, and relevant documents go unfound. This blog by @tuanacelik walks through how to set up LlamaParse for a legal discovery use-case: handling difficult scans with vision models, surfacing image and chart content, and using custom parsing instructions to guide output for predictable document patterns. The quality of everything downstream depends on what happened at ingestion. Worth getting right. Read the full blog here: https://t.co/MkUWjaJzSm

Media 1

📊 Media Metadata

{
  "media": [
    {
      "type": "photo",
      "url": "https://crmoxkoizveukayfjuyo.supabase.co/storage/v1/object/public/media/posts/2036079833000915272/media_0.png",
      "filename": "media_0.png"
    }
  ],
  "processed_at": "2026-03-23T14:01:11.507459",
  "pipeline_version": "2.0"
}

🔧 Raw API Response

{
  "type": "tweet",
  "id": "2036079833000915272",
  "url": "https://x.com/llama_index/status/2036079833000915272",
  "twitterUrl": "https://twitter.com/llama_index/status/2036079833000915272",
  "text": "If you've ever worked in or around legal, you know that discovery is where document parsing really gets stress-tested.\n\nLow-resolution scans. Black and white images. Handwritten annotations. Charts buried in reports. Files that are technically PDFs but practically unreadable. And hundreds of thousands of them.\n\nTraditional OCR tools struggle with degraded scans, and anything visual (photographs, slide decks, tables) falls through the cracks entirely. That means your search index is noisy, your recall suffers, and relevant documents go unfound.\n\nThis blog by @tuanacelik walks through how to set up LlamaParse for a legal discovery use-case: handling difficult scans with vision models, surfacing image and chart content, and using custom parsing instructions to guide output for predictable document patterns.\n\nThe quality of everything downstream depends on what happened at ingestion. Worth getting right. Read the full blog here:\n\nhttps://t.co/MkUWjaJzSm",
  "source": "Twitter for iPhone",
  "retweetCount": 0,
  "replyCount": 0,
  "likeCount": 1,
  "quoteCount": 0,
  "viewCount": 119,
  "createdAt": "Mon Mar 23 13:57:15 +0000 2026",
  "lang": "en",
  "bookmarkCount": 2,
  "isReply": false,
  "inReplyToId": null,
  "conversationId": "2036079833000915272",
  "displayTextRange": [
    0,
    280
  ],
  "inReplyToUserId": null,
  "inReplyToUsername": null,
  "author": {
    "type": "user",
    "userName": "llama_index",
    "url": "https://x.com/llama_index",
    "twitterUrl": "https://twitter.com/llama_index",
    "id": "1604278358296055808",
    "name": "LlamaIndex 🦙",
    "isVerified": false,
    "isBlueVerified": true,
    "verifiedType": "Business",
    "profilePicture": "https://pbs.twimg.com/profile_images/1967920417760251904/0ytfduMQ_normal.png",
    "coverPicture": "https://pbs.twimg.com/profile_banners/1604278358296055808/1770092126",
    "description": "AI Agents for document OCR + workflows\n\nLlamaParse: https://t.co/yQGTiRSfFL\nDocs: https://t.co/us6GCS14vD",
    "location": "",
    "followers": 110873,
    "following": 32,
    "status": "",
    "canDm": false,
    "canMediaTag": true,
    "createdAt": "Sun Dec 18 00:52:44 +0000 2022",
    "entities": {
      "description": {
        "urls": [
          {
            "display_url": "cloud.llamaindex.ai",
            "expanded_url": "https://cloud.llamaindex.ai/",
            "indices": [
              52,
              75
            ],
            "url": "https://t.co/yQGTiRSfFL"
          },
          {
            "display_url": "developers.llamaindex.ai/python/cloud/",
            "expanded_url": "https://developers.llamaindex.ai/python/cloud/",
            "indices": [
              82,
              105
            ],
            "url": "https://t.co/us6GCS14vD"
          }
        ]
      },
      "url": {
        "urls": [
          {
            "display_url": "llamaindex.ai",
            "expanded_url": "https://www.llamaindex.ai/",
            "indices": [
              0,
              23
            ],
            "url": "https://t.co/epzefqPT9Z"
          }
        ]
      }
    },
    "fastFollowersCount": 0,
    "favouritesCount": 1511,
    "hasCustomTimelines": false,
    "isTranslator": false,
    "mediaCount": 1843,
    "statusesCount": 3778,
    "withheldInCountries": [],
    "affiliatesHighlightedLabel": {},
    "possiblySensitive": false,
    "pinnedTweetIds": [
      "2029767312195117278"
    ],
    "profile_bio": {},
    "isAutomated": false,
    "automatedBy": null
  },
  "extendedEntities": {},
  "card": {
    "binding_values": [
      {
        "key": "photo_image_full_size_large",
        "value": {
          "image_value": {
            "height": 419,
            "url": "https://pbs.twimg.com/card_img/2036079835957825536/u4R9miE_?format=jpg&name=800x419",
            "width": 800
          },
          "type": "IMAGE"
        }
      },
      {
        "key": "thumbnail_image",
        "value": {
          "image_value": {
            "height": 150,
            "url": "https://pbs.twimg.com/card_img/2036079835957825536/u4R9miE_?format=jpg&name=280x150",
            "width": 266
          },
          "type": "IMAGE"
        }
      },
      {
        "key": "description",
        "value": {
          "string_value": "LlamaIndex is a simple, flexible framework for building knowledge assistants using LLMs connected to your enterprise data.",
          "type": "STRING"
        }
      },
      {
        "key": "domain",
        "value": {
          "string_value": "www.llamaindex.ai",
          "type": "STRING"
        }
      },
      {
        "key": "thumbnail_image_large",
        "value": {
          "image_value": {
            "height": 320,
            "url": "https://pbs.twimg.com/card_img/2036079835957825536/u4R9miE_?format=jpg&name=800x320_1",
            "width": 568
          },
          "type": "IMAGE"
        }
      },
      {
        "key": "summary_photo_image_small",
        "value": {
          "image_value": {
            "height": 202,
            "url": "https://pbs.twimg.com/card_img/2036079835957825536/u4R9miE_?format=jpg&name=386x202",
            "width": 386
          },
          "type": "IMAGE"
        }
      },
      {
        "key": "thumbnail_image_original",
        "value": {
          "image_value": {
            "height": 1352,
            "url": "https://pbs.twimg.com/card_img/2036079835957825536/u4R9miE_?format=jpg&name=orig",
            "width": 2400
          },
          "type": "IMAGE"
        }
      },
      {
        "key": "photo_image_full_size_small",
        "value": {
          "image_value": {
            "height": 202,
            "url": "https://pbs.twimg.com/card_img/2036079835957825536/u4R9miE_?format=jpg&name=386x202",
            "width": 386
          },
          "type": "IMAGE"
        }
      },
      {
        "key": "summary_photo_image_large",
        "value": {
          "image_value": {
            "height": 419,
            "url": "https://pbs.twimg.com/card_img/2036079835957825536/u4R9miE_?format=jpg&name=800x419",
            "width": 800
          },
          "type": "IMAGE"
        }
      },
      {
        "key": "thumbnail_image_small",
        "value": {
          "image_value": {
            "height": 81,
            "url": "https://pbs.twimg.com/card_img/2036079835957825536/u4R9miE_?format=jpg&name=144x144",
            "width": 144
          },
          "type": "IMAGE"
        }
      },
      {
        "key": "thumbnail_image_x_large",
        "value": {
          "image_value": {
            "height": 1154,
            "url": "https://pbs.twimg.com/card_img/2036079835957825536/u4R9miE_?format=png&name=2048x2048_2_exp",
            "width": 2048
          },
          "type": "IMAGE"
        }
      },
      {
        "key": "photo_image_full_size_original",
        "value": {
          "image_value": {
            "height": 1352,
            "url": "https://pbs.twimg.com/card_img/2036079835957825536/u4R9miE_?format=jpg&name=orig",
            "width": 2400
          },
          "type": "IMAGE"
        }
      },
      {
        "key": "vanity_url",
        "value": {
          "scribe_key": "vanity_url",
          "string_value": "llamaindex.ai",
          "type": "STRING"
        }
      },
      {
        "key": "photo_image_full_size",
        "value": {
          "image_value": {
            "height": 314,
            "url": "https://pbs.twimg.com/card_img/2036079835957825536/u4R9miE_?format=jpg&name=600x314",
            "width": 600
          },
          "type": "IMAGE"
        }
      },
      {
        "key": "thumbnail_image_color",
        "value": {
          "image_color_value": {
            "palette": [
              {
                "percentage": 74.57,
                "rgb": {
                  "blue": 224,
                  "green": 221,
                  "red": 236
                }
              },
              {
                "percentage": 7.89,
                "rgb": {
                  "blue": 233,
                  "green": 204,
                  "red": 224
                }
              },
              {
                "percentage": 6.78,
                "rgb": {
                  "blue": 65,
                  "green": 70,
                  "red": 82
                }
              },
              {
                "percentage": 2.9,
                "rgb": {
                  "blue": 177,
                  "green": 181,
                  "red": 247
                }
              },
              {
                "percentage": 0.51,
                "rgb": {
                  "blue": 253,
                  "green": 246,
                  "red": 216
                }
              }
            ]
          },
          "type": "IMAGE_COLOR"
        }
      },
      {
        "key": "title",
        "value": {
          "string_value": "Parsing the Unreadable: How LlamaParse Handles Legal Discovery Documents",
          "type": "STRING"
        }
      },
      {
        "key": "summary_photo_image_color",
        "value": {
          "image_color_value": {
            "palette": [
              {
                "percentage": 74.57,
                "rgb": {
                  "blue": 224,
                  "green": 221,
                  "red": 236
                }
              },
              {
                "percentage": 7.89,
                "rgb": {
                  "blue": 233,
                  "green": 204,
                  "red": 224
                }
              },
              {
                "percentage": 6.78,
                "rgb": {
                  "blue": 65,
                  "green": 70,
                  "red": 82
                }
              },
              {
                "percentage": 2.9,
                "rgb": {
                  "blue": 177,
                  "green": 181,
                  "red": 247
                }
              },
              {
                "percentage": 0.51,
                "rgb": {
                  "blue": 253,
                  "green": 246,
                  "red": 216
                }
              }
            ]
          },
          "type": "IMAGE_COLOR"
        }
      },
      {
        "key": "summary_photo_image_x_large",
        "value": {
          "image_value": {
            "height": 1154,
            "url": "https://pbs.twimg.com/card_img/2036079835957825536/u4R9miE_?format=png&name=2048x2048_2_exp",
            "width": 2048
          },
          "type": "IMAGE"
        }
      },
      {
        "key": "summary_photo_image",
        "value": {
          "image_value": {
            "height": 314,
            "url": "https://pbs.twimg.com/card_img/2036079835957825536/u4R9miE_?format=jpg&name=600x314",
            "width": 600
          },
          "type": "IMAGE"
        }
      },
      {
        "key": "photo_image_full_size_color",
        "value": {
          "image_color_value": {
            "palette": [
              {
                "percentage": 74.57,
                "rgb": {
                  "blue": 224,
                  "green": 221,
                  "red": 236
                }
              },
              {
                "percentage": 7.89,
                "rgb": {
                  "blue": 233,
                  "green": 204,
                  "red": 224
                }
              },
              {
                "percentage": 6.78,
                "rgb": {
                  "blue": 65,
                  "green": 70,
                  "red": 82
                }
              },
              {
                "percentage": 2.9,
                "rgb": {
                  "blue": 177,
                  "green": 181,
                  "red": 247
                }
              },
              {
                "percentage": 0.51,
                "rgb": {
                  "blue": 253,
                  "green": 246,
                  "red": 216
                }
              }
            ]
          },
          "type": "IMAGE_COLOR"
        }
      },
      {
        "key": "photo_image_full_size_x_large",
        "value": {
          "image_value": {
            "height": 1154,
            "url": "https://pbs.twimg.com/card_img/2036079835957825536/u4R9miE_?format=png&name=2048x2048_2_exp",
            "width": 2048
          },
          "type": "IMAGE"
        }
      },
      {
        "key": "card_url",
        "value": {
          "scribe_key": "card_url",
          "string_value": "https://t.co/MkUWjaJzSm",
          "type": "STRING"
        }
      },
      {
        "key": "summary_photo_image_original",
        "value": {
          "image_value": {
            "height": 1352,
            "url": "https://pbs.twimg.com/card_img/2036079835957825536/u4R9miE_?format=jpg&name=orig",
            "width": 2400
          },
          "type": "IMAGE"
        }
      }
    ],
    "card_platform": {
      "platform": {
        "audience": {
          "name": "production"
        },
        "device": {
          "name": "Swift",
          "version": "12"
        }
      }
    },
    "name": "summary_large_image",
    "url": "https://t.co/MkUWjaJzSm",
    "user_refs_results": []
  },
  "place": {},
  "entities": {
    "hashtags": [],
    "symbols": [],
    "urls": [
      {
        "display_url": "llamaindex.ai/blog/parsing-t…",
        "expanded_url": "https://www.llamaindex.ai/blog/parsing-the-unreadable-how-llamaparse-handles-legal-discovery-documents",
        "indices": [
          940,
          963
        ],
        "url": "https://t.co/MkUWjaJzSm"
      }
    ],
    "user_mentions": [
      {
        "id_str": "209600624",
        "indices": [
          564,
          575
        ],
        "name": "Tuana",
        "screen_name": "tuanacelik"
      }
    ]
  },
  "quoted_tweet": null,
  "retweeted_tweet": null,
  "article": null
}