Misora Sugiyama 杉山 未空
Menu

Research profile

Misora Sugiyama

杉山 未空

Computer Vision · Image & Video Generation

Portrait of Misora Sugiyama

01 — Profile

Misora Sugiyama

杉山 未空

  • NYU Tandon School of Engineering
  • LIMIT Lab
  • cvpaper.challenge

Misora Sugiyama researches computer vision and generative modeling, with a focus on image and video generation and image-to-image translation.

02 — Publications

Selected publications

Two public arXiv preprints spanning computer vision, generative modeling, image editing, and visual reasoning.

01

arXiv preprint · cs.CV

Image-Space Rule Discovery

Misora Sugiyama, Toya Oyama, Hirokatsu Kataoka

Introduces WISRD, a benchmark for testing whether image-editing models can interpret visual instructions, infer rules, and complete worksheet-style tasks directly in image space.

Image EditingVisual ReasoningBenchmark
View on arXiv
02

arXiv · ICCV 2025 Workshop

Simple Visual Artifact Detection in Sora-Generated Videos

Misora Sugiyama, Hirokatsu Kataoka

A multi-label framework for detecting four common types of visual artifacts in Sora-generated videos, achieving 94.14% average classification accuracy with ResNet-50.

Video GenerationArtifact DetectionMulti-label Classification
View on arXiv

03 — What's new?

What's new?

“Image-Space Rule Discovery” is now available on arXiv.

View ↗

“Simple Visual Artifact Detection in Sora-Generated Videos” is now available on arXiv.

View ↗

The work on visual artifact detection in generated videos was presented through an ICCV 2025 Workshop.

04 — Projects

Research projects

01

Computer Vision

Learning representations that capture objects, structure, context, and change in visual data.

RepresentationUnderstanding
02

Image & Video Generation

Studying generative models for coherent, controllable, and expressive visual synthesis.

DiffusionTemporal coherence
03

Image-to-Image Models

Translating images across domains while preserving the content and intent that matter.

TranslationControl

05 — Bio

Bio.

My work sits at the intersection of computer vision and generative AI. I am interested in models that do more than create convincing pixels—models that understand structure, preserve intent, and translate visual information across domains.

Current interests include image and video generation, image-to-image models, controllable synthesis, and the representations that connect perception with generation.

06 — Affiliations

Affiliations

NYU

NYU Tandon School of Engineering

Academic affiliation

LIM

LIMIT Lab

Research community

CV

cvpaper.challenge

Computer vision community

07 — Contact

Contact

For research inquiries, talks, or potential collaborations, feel free to get in touch.

Email Misora misorasugiyama.eca@gmail.com