Modified Date: Invitation to join 2021 Spring School ‘CVML Short Course – Computer Vision for Autonomous Systems’, 26-27th May 2021 (6th edition)

you are welcomed to register in this CVML Short e-course on ‘Computer Vision for Autonomous Systems’, 2627th May 2021: https://icarus.csd.auth.gr/spring-cvml-short-course-computer-vision-for-autonomous-systems/

 

It will take place as a two-day e-course (due to COVID-19 circumstances), hosted by the Aristotle University of Thessaloniki (AUTH), Thessaloniki, Greece, providing a series of live lectures delivered through a tele-education platform. They will be complemented with on-line video recorded lectures and lecture pdfs, to facilitate international participants having time difference issues and to enable you to study at own pace.  You can also self-assess your knowledge, by filling appropriate questionnaires (one per lecture). You will be provided programming exercises to improve your programming skills.

It is part of the very successful CVML short course series that took place in the last three years.

 

The short e-course consists of 16 1-hour live lectures organized in two Parts (1 Part per day):

Part A lectures (8 hours) provide an in-depth presentation of  2D and 3D Computer Vision theory and applications in the above-mentioned diverse domains, primarily for semantic 3D world modeling and localization. Computer Vision starts with a detailed presentation of digital image/video fundamentals and image acquisition and camera geometry, including camera calibration. Then, two lectures on: a) Stereo and Multiview imaging and b) Structure from motion will provide the theoretical and algorithmic tools to recover 3D world models from images. They will be used on Localization and mapping that is of primary importance in Autonomous Systems and Robotic perception. This is complemented by Neural techniques for recovering depth information and 3D world modeling, even from monocular images. Deep semantic image segmentation will conclude this part, by providing DNN methods both to label and segment regions, e.g., roads and targets, e.g., cars, pedestrians.

Part B lectures (8 hours) will start with an overview of Autonomous Systems Sensors. Then , it will provide an in-depth presentation of Computer Vision theory and applications in autonomous systems, particularly as related to target detection, tracking and object pose estimation. Applications will be presented for Multiple Drone Systems Autonomous Car Vision and Autonomous Marine Surface Vessels. Finally, CVML programming tools (e.g., DNN frameworks, BLAS/cuBLAS, DNN and CV libraries) are overviewed, as they allow fast application of all the above knowledge in almost any application domain.

 

Course lectures

Part A: Computer Vision (first day, 8 lectures)
  1. Digital images and videos
  2. Image Acquisition. Camera Geometry
  3. Stereo and Multiview Imaging
  4. Structure from Motion
  5. 3D Robot Localization and Mapping
  6. Neural 3D world modeling
  7. Image/Point cloud registration
  8. Deep semantic image segmentation
Part B: Autonomous Systems (second day, 8 lectures)
  1. Autonomous Systems Sensors
  2. Deep object detection
  3. Object Tracking
  4. Object Pose Estimation
  5. Multiple Drone Systems
  6. Autonomous Car Vision
  7. Autonomous Surface Vessels
  8. CVML Software Development Tools

 

Though independent, the attendees of this short e-course will greatly benefit by attending the CVML Short e-course on ‘Machine Learning and Deep Neural Networks’ 27-28th April 2021:

http://icarus.csd.auth.gr/spring-cvml-short-course-machine-learning-and-deep-neural-networks/

 

You can use the following link for course registration:

http://icarus.csd.auth.gr/spring-cvml-short-course-computer-vision-for-autonomous-systems/

 

Lecture topics, sample lecture ppts and videos, self-assessment questionnaires and programming exercises can be found therein.

For questions, please contact: Ioanna Koroni <koroniioanna@csd.auth.gr>

 

The short course is organized by Prof. I. Pitas, IEEE and EURASIP fellow, Chair of the IEEE SPS Autonomous Systems Initiative, Director of the Artificial Intelligence and Information analysis Lab (AIIA Lab), Aristotle University of Thessaloniki, Greece, Coordinator of the European Horizon2020 R&D project Multidrone. He is ranked 249-top Computer Science and Electronics scientist internationally by Guide2research (2018). He is head of the EC funded AI doctoral school of Horizon2020 EU funded R&D project AI4Media (1 of the 4 in Europe). He has 32200+ citations to his work and h-index 85+.

 

AUTH is ranked 153/182 internationally in Computer Science/Engineering, respectively, in USNews ranking.

 

Relevant links:

1) Prof. I. Pitas:

https://scholar.google.gr/citations?user=lWmGADwAAAAJ&hl=el

2) Horizon2020 EU funded R&D project Aerial-Core: https://aerial-core.eu/

3) Horizon2020 EU funded R&D project Multidrone: https://multidrone.eu/

4) Horizon2020 EU funded R&D project AI4Media: https://ai4media.eu/

5) AIIA Lab: https://aiia.csd.auth.gr/

 

Sincerely yours

Prof. I. Pitas

Director of the Artificial Intelligence and Information analysis Lab (AIIA Lab)

Aristotle University of Thessaloniki, Greece

 

Special issue on Causal Discovery and Causality-Inspired ML at IEEE TNNLS

 

Special issue in the IEEE Transactions on Neural Networks and Learning Systems (IEEE TNNLS) on Causal Discovery and Causality-Inspired Machine Learning

 

Submission deadline ** October 22, 2021 **. 

 

Call for papers in a special session: “Analyzing nondestructive evaluation data. Is it automated yet?” – ISPA2021

12th International Symposium on Image and Signal Processing and Analysis (ISPA2021)

https://www.isispa.org/

13-15th September 2021, Zagreb, Croatia

Special session Announcement

We are organizing a special session on automated analysis of the nondestructive evaluation data at the 12th International Symposium on Image and Signal Processing and Analysis (ISPA 2021). The symposium will be held in Zagreb, Croatia, on September 13-15, 2021 in a hybrid format so both physical and virtual attendance is possible. Depending on the situation with the Covid-19 pandemic the conference may shift to an online-only format.

We are inviting you to submit a paper for this special session. You may find a short description and the technical scope of this session on the following URL: 

If you are interested, please submit your paper by the 31st of May following the instructions on our official conference website. Note that this deadline might be extended in case we receive a number of requests from the authors who are not able to meet the current deadline. Feel free to share this invitation with your colleagues.


Special Session chairs

Luka Posilović, University of Zagreb, Croatia

Duje Medak, University of Zagreb, Croatia

50 JAIIO – Fecha limite llamado a presentacion de trabajos

¡FALTAN DOS SEMANAS PARA EL CIERRE DE RECEPCIÓN DE TRABAJOS!

Estimado/a,

Le recordamos que el próximo 28 de Mayo es la fecha limite para la
presentación de trabajos de las 50 JAIIO, Jornadas Argentinas de
Informática que se llevarán a cabo en un formato virtual del 18 al
29 Octubre y son co-organizadas por INTA, Instituto Nacional de
Tecnología Agropecuaria.

Le agradecemos la difusión que pueda darle!

Saludos cordiales.

�¡Seguinos en nuestras redes sociales para enterarte de más
novedades sobre las JAIIO!!

<http://www.sadio.org.ar/lists/lt.php?id=Y00NCQQDAAYLD0VTVlEOSVQPU1xX>

<http://www.sadio.org.ar/lists/lt.php?id=Y00NCQQDAAYLDkVTVlEOSVQPU1xX>

<http://www.sadio.org.ar/lists/lt.php?id=Y00NCQQDAAYMB0VTVlEOSVQPU1xX>

<http://www.sadio.org.ar/lists/lt.php?id=Y00NCQQDAAYMBkVTVlEOSVQPU1xX>

SIMPOSIOS

AGRANDA – Simposio Argentino de Ciencia de Datos y GRANdes DAtos
<http://www.sadio.org.ar/lists/lt.php?id=Y00NCQQDAAYMBUVTVlEOSVQPU1xX>

ASAI – Simposio Argentino de Inteligencia Artificial
<http://www.sadio.org.ar/lists/lt.php?id=Y00NCQQDAAYMBEVTVlEOSVQPU1xX>

ASSE – Simposio Argentino de Ingeniería de Software
<http://www.sadio.org.ar/lists/lt.php?id=Y00NCQQDAAYMA0VTVlEOSVQPU1xX>

CAI – Congreso Argentino de AgroInformática
<http://www.sadio.org.ar/lists/lt.php?id=Y00NCQQDAAYMAkVTVlEOSVQPU1xX>

CAIS – Congreso Argentino de Informática y Salud
<http://www.sadio.org.ar/lists/lt.php?id=Y00NCQQDAAYMAUVTVlEOSVQPU1xX>

EST – Concurso de Trabajos Estudiantiles
<http://www.sadio.org.ar/lists/lt.php?id=Y00NCQQDAAYMAEVTVlEOSVQPU1xX>

IETF Day – Taller del Grupo de Trabajo de Ingeniería de
Internet/Argentina
<http://www.sadio.org.ar/lists/lt.php?id=Y00NCQQDAAYMD0VTVlEOSVQPU1xX>

JUI – Jornadas de Vinculación Universidad – Industria
<http://www.sadio.org.ar/lists/lt.php?id=Y00NCQQDAAYMDkVTVlEOSVQPU1xX>

SAEI – Simposio Argentino de Educación en Informática
<http://www.sadio.org.ar/lists/lt.php?id=Y00NCQQDAAYNB0VTVlEOSVQPU1xX>

SAHTI – Simposio Argentino de Historia, Tecnologías e Informática
<http://www.sadio.org.ar/lists/lt.php?id=Y00NCQQDAAYNBkVTVlEOSVQPU1xX>

SAIV – Simposio Argentino de Imágenes y Visión
<http://www.sadio.org.ar/lists/lt.php?id=Y00NCQQDAAYNBUVTVlEOSVQPU1xX>

SID – Simposio Argentino de Informática y Derecho
<http://www.sadio.org.ar/lists/lt.php?id=Y00NCQQDAAYNBEVTVlEOSVQPU1xX>

SIE – Simposio de Informática en el Estado
<http://www.sadio.org.ar/lists/lt.php?id=Y00NCQQDAAYNA0VTVlEOSVQPU1xX>

SIIIO – Simposio Argentino de Informática Industrial e
Investigación Operativa
<http://www.sadio.org.ar/lists/lt.php?id=Y00NCQQDAAYNAkVTVlEOSVQPU1xX>

STS – Simposio Argentino sobre Tecnología y Sociedad
<http://www.sadio.org.ar/lists/lt.php?id=Y00NCQQDAAYNAUVTVlEOSVQPU1xX>

FECHAS IMPORTANTES

 Cierre de Recepción de Trabajos: 28 de mayo 2021
 Notificación de Trabajos Aceptados: 6 de agosto 2021
 Recepción de Trabajos “Camera Ready”: 20 de agosto de 2021
 Realización de la conferencia: 18 al 29 de Octubre de 2021

The First Continual Semi-Supervised Learning Challenge @ IJCAI 2021


The First Continual Semi-Supervised Learning Challenge

Call for participation

The Challenge is organised as part of the upcoming IJCAI 2021 First International Workshop on Continual Semi-Supervised Learning

https://sites.google.com/view/sscl-workshop-ijcai-2021/

Aim of the Workshop

Whereas continual learning has recently attracted much attention in the machine learning community, the focus has been mainly on preventing the model updated in the light of new data from ‘catastrophically forgetting’ its initial knowledge and abilities. This, however, is in stark contrast with common real-world situations in which an initial model is trained using limited data, only to be later deployed without any additional supervision. In these scenarios the goal is for the model to be incrementally updated using the new (unlabelled) data, in order to adapt to a target domain continually shifting over time.

The aim of this workshop is to formalise this new continual semi-supervised learning paradigm, and to introduce it to the machine learning community in order to mobilise effort in this direction. We present the first two benchmark datasets for this problem, derived from significant computer vision scenarios, and propose the first Continual Semi-Supervised Learning Challenges to the research community.

Problem Statement

In semi-supervised continual learning, an initial training batch of data points annotated with ground truth (class labels for classification problems, or vectors of target values for regression ones) is available and can be used to train an initial model. Then, however, the model is incrementally updated by exploiting the information provided by a time series of unlabelled data points, each of which is generated by a data generating process (modelled, as typically assumed, by a probability distribution) which may vary with time, without any artificial subdivision into ‘tasks’.

Challenges

We propose both a continual activity recognition (CAR) challenge and a continual crowd counting (CCC) challenge.

https://sites.google.com/view/sscl-workshop-ijcai-2021/challenges

In the former, the aim is to devise a learning mechanism for updating a baseline action recognition method (working at frame level) based on a data stream of video frames, of which only the initial fraction is labelled (a classification problem).

In the latter, the learning mechanism is applied to a baseline crowd counting method, also working on a frame-by-frame basis, and exploits a data stream of video frames of which only an initial fraction come with ground truth attached in the form of a density map (a regression problem).

Benchmark Datasets

As a benchmark for the continual activity recognition challenge we have created a Continual Activity Recognition (CAR) dataset, derived from a fraction of the MEVA (Multiview Extended Video with Activities) activity detection dataset (https://mevadata.org/). We selected a suitable set of 8 activity classes from the original list of 37, and annotated each frame in 15 video sequences, each composed by 3 clips originally from MEVA, with a single class label.

Our CAR benchmark is thus composed of 15 sequences, broken down into three groups:

· Five 15-minute-long sequences formed by three original videos which are contiguous.

· Five 15-minute-long sequences formed by three videos separated by a short time interval (5-20 minutes).

· Five 15-minute-long sequences formed by three original videos separated by a long interval of time (hours or even days).

Each of these three evaluation settings is designed to simulate a different mix of continuous and discrete dynamics of the domain distribution.

The raw video sequences are directly accessible from the Challenge website.

Our CCC benchmark is composed of 3 sequences, taken from existing crowd counting datasets:

· A single 2,000 frame sequence from the Mall dataset.

· A single 2,000-frame sequence from the UCSD dataset.

· A 750-frame sequence from the Fudan-ShanghaiTech (FDST) dataset, composed of 5 clips portraying the same scene each 150 frames long.

Ground truth

The ground truth for the CAR challenges (in the form of one activity label per frame) was created by us, after selecting a subset of 8 activity classes and revising the original annotation for the 45 video clips we selected for inclusion.

The ground truth for the CCC challenges (in the form of a density map for each frame) was generated by us for all three datasets following the annotation protocol described in

https://github.com/svishwa/crowdcount-mcnn

The ground truth for both challenges will be released on the Challenge web site according to the following schedule:

·        Training and validation fold release: May 5 2021

·        Test fold release: June 30 2021

·        Submission of results: July 15 2021

·        Announcement of results: July 31 2021

·        Challenge event @ workshop: August 21-23 2021

Tasks

For each challenge we propose two separate tasks, incremental and absolute.

CAR-A: The goal is to achieve the best average performance across all the unlabelled test portion of the 15 sequences in the CAR dataset. The choice of the baseline action recognition model is left to the participants.

CAR-I: The goal here is to achieve the best performance improvement over the baseline (supervised) model, measured on the unlabelled test data stream, on average over the 15 sequences. The baseline recognition model is set by us (see Baselines below).

CCC-A: Seeks the best average performance over the unlabelled test portion of the 3 sequences of the CCC dataset.  The choice of the baseline crowd counting model is left to the participants.

CCC-I: Seeks the best performance improvement over the baseline, measured on the unlabelled test portion of the data stream, on average over the 3 sequences of CCC. The baseline crowd counting model is set by us (see Baselines below).

Protocol for Incremental Training and Testing

Following from the problem definition, once a model is fine-tuned on the supervised portion of a data stream it is subsequently both incrementally updated using the unlabelled portion of the same data stream and tested there, using the available ground truth (encapsulated in an evaluation script).

Importantly, incremental training and testing must happen independently for each sequence, as we intent to simulate real-world scenarios in which a smart device with continual learning capability can only learn from its own data stream after deployment.

The two challenges differ in the sense that, whereas in CAR the baseline activity recognition model is initially fine-tuned using the supervised folds of all the 15 available sequences jointly (since each sequence only portrays a subset of the 8 activity classes), in CCC the supervised fine-tuning happens sequence by sequence (because of the disparate nature of the videos captured in different settings).

Split into Training, Validation and Test

The data for the challenges are released in two Stages:

1.       We first release the supervised portion of each data stream, together with a portion of the unlabelled data stream to use for the validation of the semi-supervised continual learning approach proposed by the participants.

2.       The remaining portion of the unlabelled data stream for each sequence in the dataset is released at a later stage to be used for the testing of the proposed approach.

Consequently, each data stream (sequence) in our benchmarks is divided into a supervised fold (S), a validation fold (V) and a test fold (T).

For the CAR challenge, the supervised fold for each sequence coincides with the first 5-minute video, the validation fold with the second 5-minute video, and the test fold with the third 5-minute video.

For the CCC challenge we distinguish two cases. For the 2,000-frame sequences from either the UCSD or the Mall dataset, S is formed by the first 400 images, V by the following 800 images, and T by the remaining 800 images. For the 750-frame sequence from the FDST dataset, S is the set of the first 150 images, V the set of the following 300 images, and T the set of the remaining 300 images.

Evaluation

Participants will be able to evaluate the performance of their method(s) on both the incremental and the absolute versions of the challenges on eval.ai.

In Stage 1 participants will, for each task (CAR-A, CAR-I, CCC-A, CCC-I), submit their predictions as generated on the validation folds and get the evaluation metric in return, in order to get a feel of how well their method(s) work. In Stage 2 they will submit the predictions generated on the test folds which will be used for the final ranking.

A separate ranking will be produced for each of the tasks.

For each of the challenge stages and each task the maximum number of submissions through the EvalAI platform is capped at 50, with an additional constraint of 5 submissions per day.

Detailed instructions about how to download the data and submit your predictions for evaluation at both validation and test time, for all four tasks, are provided here:

https://sites.google.com/view/sscl-workshop-ijcai-2021/challenges

Baselines

Baselines are provided regarding the initial action recognition model to be used in CAR-A, the base crowd counter to be used in CCC-A, and the semi-supervised incremental learning process itself.

Baseline activity recognition model

As our activity recognition baseline model we adopted the recent EfficientNet[1] network (model EfficientNet-B5), pre-trained on the large-scale ImageNet dataset. Detailed information about its implementation, along with pre-trained models, can be found on Github

https://github.com/lukemelas/EfficientNet-PyTorch

and is easily downloadable using the Python command “pip” (pip install efficientnet-pytorch).

Note that the performance of the baseline activity model is rather poor on our challenge videos, as relevant activities occupy only a small fraction of the duration of the videos. This leaves much room for improvement while properly representing the level of challenge real-world data poses.

Baseline crowd counter

As baseline crowd counting model we selected the Multi-Column Convolutional Neural Network (MCNN)[2]. Its implementation, along with pre-trained models, can also be found on Github

https://github.com/svishwa/crowdcount-mcnn

This network is implemented using PyTorch. Pre-trained models are available for both the ShanghaiTech A and the ShanghaiTech B datasets. For this Challenge we chose to adopt the ShanghaiTechB pre-trained model.

Baseline incremental learning approach

Finally, our baseline for incremental learning from unlabelled data stream is based on a vanilla (batch) self-training approach.

For each sequence, the unlabelled data stream (without distinction between validation and test folds) is partitioned into a number of sub-folds. Each sub-fold spans 1 minute in the CAR challenges, so that each unlabelled sequence is split into 10 sub-folds. Sub-folds span 100 frames in the CCC challenges, so that the UCSD and MALL (unlabelled) data streams are decomposed into 16 sub-folds whereas the FDST sequence only contains 6 sub-folds.

Starting with the model initially fine-tuned on the supervised portion of the data stream, self-training is iteratively applied in a batch fashion to each sub-fold. The predictions generated by the model obtained after convergence upon a sub-fold are the baseline predictions for the current sub-fold. The output of each self-training session is used as start model for the following session.

Reproducibility

Participants will be clearly told that we reserve the right to reproduce their results and check their validity. We will certainly reproduce the results of the challenge winners.



[1] Tan, M., & Le, Q. (2019, May). EfficientNet: Rethinking model scaling for convolutional neural networks. In International Conference on Machine Learning (pp. 6105-6114). PMLR.

[2] Zhang, Yingying, et al. “Single-image crowd counting via multi-column convolutional neural network.” Proceedings of the IEEE conference on computer vision and pattern recognition. 2016. 

Design by 2b Consult