Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Kunyang Zhou

Unsupervised Domain Adaptive Lane Detection via Contextual Contrast and Aggregation

Jul 18, 2024

Kunyang Zhou, Yunjian Feng, Jun Li

Figure 1 for Unsupervised Domain Adaptive Lane Detection via Contextual Contrast and Aggregation

Figure 2 for Unsupervised Domain Adaptive Lane Detection via Contextual Contrast and Aggregation

Figure 3 for Unsupervised Domain Adaptive Lane Detection via Contextual Contrast and Aggregation

Figure 4 for Unsupervised Domain Adaptive Lane Detection via Contextual Contrast and Aggregation

Abstract:This paper focuses on two crucial issues in domain-adaptive lane detection, i.e., how to effectively learn discriminative features and transfer knowledge across domains. Existing lane detection methods usually exploit a pixel-wise cross-entropy loss to train detection models. However, the loss ignores the difference in feature representation among lanes, which leads to inefficient feature learning. On the other hand, cross-domain context dependency crucial for transferring knowledge across domains remains unexplored in existing lane detection methods. This paper proposes a method of Domain-Adaptive lane detection via Contextual Contrast and Aggregation (DACCA), consisting of two key components, i.e., cross-domain contrastive loss and domain-level feature aggregation, to realize domain-adaptive lane detection. The former can effectively differentiate feature representations among categories by taking domain-level features as positive samples. The latter fuses the domain-level and pixel-level features to strengthen cross-domain context dependency. Extensive experiments show that DACCA significantly improves the detection model's performance and outperforms existing unsupervised domain adaptive lane detection methods on six datasets, especially achieving the best performance when transferring from CULane to Tusimple (92.10% accuracy), Tusimple to CULane (41.9% F1 score), OpenLane to CULane (43.0% F1 score), and CULane to OpenLane (27.6% F1 score).

Via

Access Paper or Ask Questions

Lane2Seq: Towards Unified Lane Detection via Sequence Generation

Feb 27, 2024

Kunyang Zhou

Figure 1 for Lane2Seq: Towards Unified Lane Detection via Sequence Generation

Figure 2 for Lane2Seq: Towards Unified Lane Detection via Sequence Generation

Figure 3 for Lane2Seq: Towards Unified Lane Detection via Sequence Generation

Figure 4 for Lane2Seq: Towards Unified Lane Detection via Sequence Generation

Abstract:In this paper, we present a novel sequence generation-based framework for lane detection, called Lane2Seq. It unifies various lane detection formats by casting lane detection as a sequence generation task. This is different from previous lane detection methods, which depend on well-designed task-specific head networks and corresponding loss functions. Lane2Seq only adopts a plain transformer-based encoder-decoder architecture with a simple cross-entropy loss. Additionally, we propose a new multi-format model tuning based on reinforcement learning to incorporate the task-specific knowledge into Lane2Seq. Experimental results demonstrate that such a simple sequence generation paradigm not only unifies lane detection but also achieves competitive performance on benchmarks. For example, Lane2Seq gets 97.95\% and 97.42\% F1 score on Tusimple and LLAMAS datasets, establishing a new state-of-the-art result for two benchmarks.

* CVPR2024 acceptance

Via

Access Paper or Ask Questions

End-to-End Lane detection with One-to-Several Transformer

May 13, 2023

Kunyang Zhou, Rui Zhou

Figure 1 for End-to-End Lane detection with One-to-Several Transformer

Figure 2 for End-to-End Lane detection with One-to-Several Transformer

Figure 3 for End-to-End Lane detection with One-to-Several Transformer

Figure 4 for End-to-End Lane detection with One-to-Several Transformer

Abstract:Although lane detection methods have shown impressive performance in real-world scenarios, most of methods require post-processing which is not robust enough. Therefore, end-to-end detectors like DEtection TRansformer(DETR) have been introduced in lane detection.However, one-to-one label assignment in DETR can degrade the training efficiency due to label semantic conflicts. Besides, positional query in DETR is unable to provide explicit positional prior, making it difficult to be optimized. In this paper, we present the One-to-Several Transformer(O2SFormer). We first propose the one-to-several label assignment, which combines one-to-many and one-to-one label assignment to solve label semantic conflicts while keeping end-to-end detection. To overcome the difficulty in optimizing one-to-one assignment. We further propose the layer-wise soft label which dynamically adjusts the positive weight of positive lane anchors in different decoder layers. Finally, we design the dynamic anchor-based positional query to explore positional prior by incorporating lane anchors into positional query. Experimental results show that O2SFormer with ResNet50 backbone achieves 77.83% F1 score on CULane dataset, outperforming existing Transformer-based and CNN-based detectors. Futhermore, O2SFormer converges 12.5x faster than DETR for the ResNet18 backbone.

* code: https://github.com/zkyseu/O2SFormer

Via

Access Paper or Ask Questions