
This project is an automated video captioning agent designed for the AMD Hackathon. The pipeline processes video content by separating the task into four straightforward steps: Ingestion: The system samples frames from the video and transcribes the audio track into text using a local transcription model. Extraction: The sampled frames and audio text are sent to a Vision-Language Model, which identifies the raw, objective facts of the scene (e.g., physical objects, setting, actions, and lighting). Compression: The raw facts are grouped and simplified into a clean timeline of events, removing redundant information. Synthesis: Finally, the timeline and key frames are fed into a large language model. The model is given strict persona instructions to apply the requested styles (Formal, Sarcastic, etc.) while staying firmly grounded in the factual timeline, ensuring the captions are both stylized and visually accurate.
13 Jul 2026

Scrappy is poised to transform how the world extracts and leverages data from the web. Our journey starts with the launch of the core autonomous API generator, giving anyone—from analysts to entrepreneurs—the power to turn static websites into usable APIs in seconds. This foundation eliminates the fragility of traditional scraping and opens access to web data for a much wider audience. The next step is expansion. Scrappy will move beyond static sites to support dynamic, JavaScript-driven platforms, ensuring compatibility with the modern web. Alongside this, we will integrate multi-agent intelligence, allowing Scrappy to analyze, generate, test, and refine APIs in a collaborative and autonomous way. This evolution will create a more adaptive, resilient, and intelligent data-extraction system. As Scrappy matures, we will enhance intelligence and reliability through advanced monitoring, analytics, and enterprise-ready features. Businesses will not just receive an API—they will gain intelligent data pipelines that monitor themselves, adapt to change, and deliver consistent results over time. This will position Scrappy as both a developer-friendly tool and a robust enterprise platform. Ultimately, Scrappy is about more than scraping—it is about building the autonomous data layer of the future. A world where any website can be instantly transformed into an accessible, structured API. By lowering barriers to access, Scrappy democratizes web data, empowering individuals and organizations to innovate faster, make smarter decisions, and unlock new opportunities. The vision is bold but simple: launch strong, expand capabilities, enhance intelligence, and pave the way toward a web where information flows freely as APIs. Scrappy is not just solving a problem—it’s shaping the future of data accessibility.
24 Aug 2025