Sitemap

A list of all the posts and pages found on the site. For you robots out there, there is an XML version available for digesting as well.

Pages

Posts

Future Blog Post

less than 1 minute read

Published:

This post will show up by default. To disable scheduling of future posts, edit config.yml and set future: false.

Blog Post number 4

less than 1 minute read

Published:

This is a sample blog post. Lorem ipsum I can’t remember the rest of lorem ipsum and don’t have an internet connection right now. Testing testing testing this blog post. Blog posts are cool.

Blog Post number 3

less than 1 minute read

Published:

This is a sample blog post. Lorem ipsum I can’t remember the rest of lorem ipsum and don’t have an internet connection right now. Testing testing testing this blog post. Blog posts are cool.

Blog Post number 2

less than 1 minute read

Published:

This is a sample blog post. Lorem ipsum I can’t remember the rest of lorem ipsum and don’t have an internet connection right now. Testing testing testing this blog post. Blog posts are cool.

Blog Post number 1

less than 1 minute read

Published:

This is a sample blog post. Lorem ipsum I can’t remember the rest of lorem ipsum and don’t have an internet connection right now. Testing testing testing this blog post. Blog posts are cool.

portfolio

publications

DIVE-Doc: Downscaling foundational Image Visual Encoder into hierarchical architecture for DocVQA

Published in ICCV (VisionDocs Workshop), 2025 | Code Repository | Demo
Oral Spotlight & Best Paper Award
Rayane Bencharef, Abderrahmane Rahiche, Mohamed Cheriet

This paper is about optimizing Visual Encoder (VE) of end-to-end DocVQA architectures. While reducing by 5x the size of the VE from a foundational to a hierarchical architecture, DIVE-Doc achieves a performance gap of 2.10 (ANLS) compared to its teacher Paligemma through a distillation training process. Moreover, this allowed to halves the visual module latency. Evaluation of the VE on downstream tasks (document classification & layout analysis) led to the conclusion that the VE seems to capture global structure of documents while semantic layout understanding seems to be achieved inside the language model decoder.

talks

teaching

Teaching experience 1

Undergraduate course, University 1, Department, 2014

This is a description of a teaching experience. You can use markdown like any other post.

Teaching experience 2

Workshop, University 1, Department, 2015

This is a description of a teaching experience. You can use markdown like any other post.