Our journal article on Recent trends in Crowd Analysis : a review published in Machine Learning with Applications (Elsevier) is now available in open access on ScienceDirect. Here is the link to it: https://t.co/h2dq4jZyaA
@dh7net Thank you for this thread! I haven't read the paper, yet, but if I understand well, the main contribution of InstructPix2Pix is to prompt diffusion models differently. Now, instructions can be written in an imperative style rather than a declarative style. Am I right?
@rainisto@alexpolymath What is the image size you did choose for your training samples ? Did you stick with the 512x512 size or as SD 2.1 generate 768² images, did you resize the training images to 768x768 pixels ?
We (@CharlottePlltr@SylvainLobry@gforestier and Camille Kurtz) are organizing a special session at JURSE 2023 about "Deep learning approaches for multi-temporal and multi-modal data processing and analysis for urban areas". https://t.co/9bwJpzLjIb
Call for papers: Workshop On Learning With Few or Without Annotated Face, Body and Gesture Data at #FG2023@IEEEBiometrics
Submission deadline extended to 09/22/2022
More info: https://t.co/sUCszByxuj
Co-organized with @DaoudiMed64, Stefano Berretti, @gforestier and @jjmweber
THREAD : Les 10 sites gratuits qui vont révolutionner votre vie de dev.
Voici ma liste ultime des outils en ligne qui m'ont économisé des journées entières de boulot ⤵️
We have an open research engineer position on virtual reality for rehabilitation movements assessment. The position is in the MSD research group of IRIMAS Institute at @UHA68, in collaboration with @GEscrivaBoulley from LISEC Lab.
#virtualreality#humanmotion#rehabilitation
Unidentified Video Objects (UVO): a new benchmark for open-world object segmentation in videos.
UVO aims to enable development and research of machine learning methods for more comprehensive video understanding.
https://t.co/wdpurlVZ9s
Interested in Deep Learning and human motion analysis ? Apply to our Master internship position in deep generative models for skeleton-based human motion. More information at https://t.co/yKwcISLdOk
#deeplearning#humanmotion#generativemodel#timeseries
🤖 BEHAVIOR: A new benchmark with the 100 household activities that represent a new challenge for embodied AI solutions.
Agents need to navigate and manipulate the simulated environment with the goal of accomplishing the activities.
https://t.co/JldfEcO6j2
🎞️ VALUE: A video-and-language understanding evaluation benchmark to test models that are generalizable to diverse tasks and domains. It's an assemblage of 11 datasets over 3 tasks: text-to-video retrieval, video question answering; and video captioning.
https://t.co/lVKrgOaOvL
VATT: Transformers for Multimodal Self-Supervised Learning from Raw Video, Audio and Text
https://t.co/MsPuanMo6D
by Hassan Akbari et al. including @YinCui1#ComputerScience#Learning
Noise-Aware Saliency Prediction for Videos with Incomplete Gaze Data
https://t.co/78EzONZpAz
by Ekta Prashnani et al. including @0razio#LossFunction#DeepLearning
Modular Interactive Video Object Segmentation: Interaction-to-Mask, Propagation and Difference-Aware Fusion
https://t.co/A4Fdza811H
by Ho Kei Cheng et al.
#ComputerVision#PatternRecognition