检索结果-内蒙古大学图书馆

您好，读者！请登录

内蒙古大学图书馆

首页
概况
党建
资源
服务
科研支持
- 论文收录引用证明
- 科技查新
知识产权
档案馆
帮助

咨询与建议

建议与咨询留下您的常用邮箱和电话号码，以便我们向您反馈解决方案和替代方法

您的常用邮箱：*

您的手机号码：*

问题描述：

当前已输入0个字，您还可以输入200个字

全部搜索
期刊论文
图书
学位论文
标准
纸本馆藏
外文资源发现
数据库导航
超星发现

高级检索

分类表

所选分类

>> <<

限定检索结果

标题

标题
作者
主题词
出版物名称
出版社
机构
学科分类号
摘要
ISBN
ISSN
基金资助
索书号

作者

作者
标题
主题词
出版物名称
出版社
机构
学科分类号
摘要
ISBN
ISSN
基金资助
索书号

文献类型

3,915 篇 会议
2 篇 期刊文献

馆藏范围

3,917 篇 电子文献
0 种 纸本馆藏

日期分布

学科分类号

3,011 篇 工学
- 2,919 篇 计算机科学与技术...
- 216 篇 软件工程
- 141 篇 机械工程
- 133 篇 光学工程
- 42 篇 生物工程
- 28 篇 信息与通信工程
- 25 篇 电气工程
- 17 篇 控制科学与工程
- 9 篇 电子科学与技术（可...
- 9 篇 化学工程与技术
- 9 篇 交通运输工程
- 8 篇 生物医学工程（可授...
- 7 篇 安全科学与工程
- 4 篇 材料科学与工程（可...
- 4 篇 建筑学
- 3 篇 土木工程
- 3 篇 农业工程
1,559 篇 医学
- 1,558 篇 临床医学
- 3 篇 基础医学(可授医学...
174 篇 理学
- 136 篇 物理学
- 43 篇 生物学
- 29 篇 数学
- 16 篇 统计学（可授理学、...
- 10 篇 化学
14 篇 管理学
- 7 篇 管理科学与工程(可...
- 7 篇 图书情报与档案管...
- 3 篇 工商管理
5 篇 法学
- 3 篇 社会学
- 2 篇 法学
2 篇 教育学
- 2 篇 教育学
2 篇 农学
1 篇 经济学

主题

2,408 篇 computer vision
1,085 篇 training
1,043 篇 pattern recognit...
805 篇 conferences
709 篇 computational mo...
543 篇 visualization
491 篇 computer archite...
446 篇 three-dimensiona...
410 篇 semantics
409 篇 benchmark testin...
383 篇 codes
331 篇 transformers
290 篇 deep learning
277 篇 feature extracti...
260 篇 neural networks
256 篇 task analysis
238 篇 shape
216 篇 image segmentati...
204 篇 measurement
202 篇 object detection

机构

72 篇 tsinghua univ pe...
59 篇 univ sci & techn...
55 篇 zhejiang univ pe...
55 篇 chinese univ hon...
51 篇 carnegie mellon ...
51 篇 peng cheng lab p...
49 篇 swiss fed inst t...
47 篇 sensetime res pe...
46 篇 shanghai ai lab ...
42 篇 univ hong kong p...
38 篇 huawei noahs ark...
35 篇 univ chinese aca...
35 篇 shanghai jiao to...
33 篇 alibaba grp peop...
32 篇 tech univ munich...
31 篇 stanford univ st...
30 篇 peking univ peop...
30 篇 swiss fed inst t...
29 篇 adobe res san jo...
29 篇 google res mount...

作者

63 篇 timofte radu
30 篇 van gool luc
19 篇 yang yi
19 篇 qiao yu
18 篇 loy chen change
16 篇 zhang lei
16 篇 radu timofte
15 篇 sun jian
14 篇 liu yang
14 篇 liu shuaicheng
14 篇 tao dacheng
13 篇 li xin
13 篇 fan haoqiang
13 篇 chen wei-ting
12 篇 luo ping
12 篇 chen dongdong
12 篇 wang kai
12 篇 wang xinchao
12 篇 torralba antonio
12 篇 ghanem bernard

语言

3,916 篇 英文
1 篇 其他

检索条件"任意字段=2022 IEEE/CVF Conference on Computer Vision and Pattern Recognition Workshops, CVPRW 2022"

共 3917 条记录，以下是51-60 订阅

全选清除本页清除全部题录导出标记到"检索档案"

详细简洁

排序：

相关度排序

相关度排序
时效性降序
时效性升序

ALINA: Advanced Line Identification and Notation Algorithm

ALINA: Advanced Line Identification and Notation Algorithm

引用

ieee/cvf conference on computer vision and pattern recognition (CVPR)

作者： Khan, Mohammed Abdul Hafeez Ganeriwala, Parth Bhattacharyya, Siddhartha Neogi, Natasha Muthalagu, Raja Florida Inst Technol Melbourne FL 32901 USA NASA Langley Res Ctr Hampton VA 23665 USA BITS Pilani Dubai Campus Dubai U Arab Emirates

ISBN: (纸本)9798350365474

Labels are the cornerstone of supervised machine learning algorithms. Most visual recognition methods are fully supervised, using bounding boxes or pixel-wise segmentations for object localization. Traditional labeling methods, such as crowd-sourcing, are prohibitive due to cost, data privacy, amount of time, and potential errors on large datasets. To address these issues, we propose a novel annotation framework, Advanced Line Identification and Notation Algorithm (ALINA), which can be used for labeling taxiway datasets that consist of different camera perspectives and variable weather attributes (sunny and cloudy). Additionally, the CIRCular threshoLd pixEl Discovery And Traversal (CIRCLEDAT) algorithm has been proposed, which is an integral step in determining the pixels corresponding to taxiway line markings. Once the pixels are identified, ALINA generates corresponding pixel coordinate annotations on the frame. Using this approach, 60,249 frames from the taxiway dataset, AssistTaxi have been labeled. To evaluate the performance, a context-based edge map (CBEM) set was generated manually based on edge features and connectivity. The detection rate after testing the annotated labels with the CBEM set was recorded as 98.45%, attesting its dependability and effectiveness.

关键词： aircraft perception annotation autonomous driving computer vision labeling line identification taxiway data

来源：评论

学校读者我要写书评

暂无评论

Classifier Guided Cluster Density Reduction for Dataset Selection

Classifier Guided Cluster Density Reduction for Dataset Sele...

引用

ieee/cvf conference on computer vision and pattern recognition (CVPR)

作者： Chang, Cheng Long, Keyu Li, Zijian Rai, Himanshu Layer 6 AI Toronto ON Canada

ISBN: (纸本)9798350365474

In this paper, we address the challenge of selecting an optimal dataset from a source pool with annotations to enhance performance on a target dataset derived from a different source. This is important in scenarios where it is hard to afford on-the-fly dataset annotation and is also the theme of the second Visual Data Understanding (VDU) Challenge. Our solution, the Classifier Guided Cluster Density Reduction (CCDR) framework, operates in two stages. Initially, we employ a filtering technique to identify images that align with the target dataset's distribution. Subsequently, we implement a graph-based cluster density reduction method, steered by a classifier that approximates the distance between the target distribution and source distribution. This classifier is trained to distinguish between images that resemble the target dataset and those that do not, facilitating the pruning process shown in Figure 1. Our approach maintains a balance between selecting pertinent images that match the target distribution and eliminating redundant ones that do not contribute to the enhancement of the detection model. We demonstrate the superiority of our method over various baselines in object detection tasks, particularly in optimizing the training set distribution on the region100 dataset. We have released our code here: https://***/ himsR/DataCVChallenge-2024/tree/main

关键词： computer vision Data Search deep learning domain Transfer

来源：评论

学校读者我要写书评

暂无评论

Domain Targeted Synthetic Plant Style Transfer using Stable Diffusion, LoRA and ControlNet

Domain Targeted Synthetic Plant Style Transfer using Stable ...

引用

ieee/cvf conference on computer vision and pattern recognition (CVPR)

作者： Hartley, Zane K. J. Lind, Rob J. Pound, Michael P. French, Andrew P. Univ Nottingham Wollaton Rd Nottingham NG8 1BB England Syngenta Jealotts Hill Int Res Ctr Warfield England

ISBN: (纸本)9798350365474

Synthetic images can help alleviate much of the cost in the creation of training data for plant phenotyping-focused AI development. Synthetic-to-real style transfer is of particular interest to users of artificial data because of the domain shift problem created by training neural networks on images generated in a digital environment. In this paper we present a pipeline for synthetic plant creation and image-to-image style transfer, with a particular interest in synthetic to real domain adaptation targeting specific real datasets. Utilizing new advances in generative AI, we employ a combination of Stable diffusion, Low Ranked Adapters (LoRA) and ControlNets to produce an advanced system of style transfer. We focus our work on the core task of leaf instance segmentation, exploring both synthetic to real style transfer as well as inter-species style transfer and find that our pipeline makes numerous improvements over CycleGAN for style transfer, and the images we produce are comparable to real images when used as training data.

关键词： Agriculture computer vision ControlNet Deep Learning Diffusion LoRA Plant Phenotyping

来源：评论

学校读者我要写书评

暂无评论

Towards Engineered Safe AI with Modular Concept Models

Towards Engineered Safe AI with Modular Concept Models

引用

ieee/cvf conference on computer vision and pattern recognition (CVPR)

作者： Heidemann, Lena Kurzidem, Iwo Monnet, Maureen Roscher, Karsten Guennemann, Stephan Fraunhofer IKS Munich Germany Tech Univ Munich Munich Germany

ISBN: (纸本)9798350365474

The inherent complexity and uncertainty of Machine Learning (ML) makes it difficult for ML-based computer vision (CV) approaches to become prevalent in safety-critical domains like autonomous driving, despite their high performance. A crucial challenge in these domains is the safety assurance of ML-based systems. To address this, recent safety standardization in the automotive domain has introduced an ML safety lifecycle following an iterative development process. While this approach facilitates safety assurance, its iterative nature requires frequent adaptation and optimization of the ML function, which might include costly retraining of the ML model and is not guaranteed to converge to a safe AI solution. In this paper, we propose a modular ML approach which allows for more efficient and targeted measures to each of the modules and process steps. Each module of the modular concept model represents one visual concept and is aggregated with the other modules' outputs into a task output. The design choices of a modular concept model can be categorized into the selection of the concept modules, the aggregation of their output and the training of the concept modules. Using the example of traffic sign classification, we present each step of the involved design choices and the corresponding targeted measures to take in an iterative development process for engineering safe AI.

关键词： computer vision Concept Models Deep Neural Networks Explainable AI Interpretable Models Machine Learning ML Safety Modular Concept Models Modular Deep Learning Safe AI Safe ML

来源：评论

学校读者我要写书评

暂无评论

ZInD-Tell: Towards Translating Indoor Panoramas into Descriptions

ZInD-Tell: Towards Translating Indoor Panoramas into Descrip...

引用

ieee/cvf conference on computer vision and pattern recognition (CVPR)

作者： Deb, Tonmoay Wang, Lichen Bessinger, Zachary Khosravan, Naji Penner, Eric Kang, Sing Bing Northwestern Univ Evanston IL 60208 USA Zillow Grp Seattle WA USA

ISBN: (纸本)9798350365474

This paper focuses on bridging the gap between natural language descriptions, 360 degrees panoramas, room shapes, and layouts/floorplans of indoor spaces. To enable new multimodal (image, geometry, language) research directions in indoor environment understanding, we propose a novel extension to the Zillow Indoor Dataset (ZInD) which we call ZInD-Tell1. We first introduce an effective technique for extracting geometric information from ZInD's raw structural data, which facilitates the generation of accurate ground truth descriptions using GPT-4. A human-in-the-loop approach is then employed to ensure the quality of these descriptions. To demonstrate the vast potential of our dataset, we introduce the ZInD-Tell benchmark, focusing on two exemplary tasks: language-based home retrieval and indoor description generation. Furthermore, we propose an end-to-end, zero-shot baseline model, ZInD-Agent, designed to process an unordered set of panorama images and generate home descriptions. ZInD-Agent outperforms naive methods in both tasks, hence, can be considered as a complement to the naive to show potential use of the data and impact of geometry. We believe this work initiates new trajectories in leveraging computer vision techniques to analyze indoor panorama images descriptively by learning the latent relation between vision, geometry, and language modalities.

关键词： computer vision

来源：评论

学校读者我要写书评

暂无评论

ViTKD: Feature-based Knowledge Distillation for vision Transformers

ViTKD: Feature-based Knowledge Distillation for Vision Trans...

引用

ieee/cvf conference on computer vision and pattern recognition (CVPR)

作者： Yang, Zhendong Li, Zhe Zeng, Ailing Li, Zexian Yu, Chun Yu, Liming Tsinghua Shenzhen Int Grad Sch Shenzhen Peoples R China Int Digital Econ Acad IDEA Shenzhen Peoples R China Chinese Acad Sci Inst Automat Beijing Peoples R China Beihang Univ Beijing Peoples R China IDEA Shenzhen Peoples R China

ISBN: (纸本)9798350365474

Knowledge Distillation (KD) has been extensively studied as a means to enhance the performance of smaller models in Convolutional Neural Networks (CNNs). Recently, the vision Transformer (ViT) has demonstrated remarkable success in various computer vision tasks, leading to an increased demand for KD in ViT. However, while logit-based KD has been applied to ViT, other feature-based KD methods for CNNs cannot be directly implemented due to the significant structure gap. In this paper, we conduct an analysis of the properties of different feature layers in ViT to identify a method for feature-based ViT distillation. Our findings reveal that both shallow and deep layers in ViT are equally important for distillation and require distinct distillation strategies. Based on these guidelines, we propose our feature-based method ViTKD, which mimics the shallow layers and generates the deep layer in the teacher. ViTKD leads to consistent and significant improvements in the students. On ImageNet-1K, we achieve performance boosts of 1.64% for DeiT-Tiny, 1.40% for DeiT-Small, and 1.70% for DeiT-Base. Downstream tasks also demonstrate the superiority of ViTKD. Additionally, ViTKD and logit-based KD are complementary and can be applied together directly, further enhancing the student's performance. Specifically, DeiT-T, S, and B achieve accuracies of 77.78%, 83.59%, and 85.41%, respectively, using this combined approach. Code is available at https:// github. com/ yzdv/cls_KD.

关键词： Convolutional neural networks

来源：评论

学校读者我要写书评

暂无评论

VLM-PL: Advanced Pseudo Labeling approach for Class Incremental Object Detection via vision-Language Model

VLM-PL: Advanced Pseudo Labeling approach for Class Incremen...

引用

ieee/cvf conference on computer vision and pattern recognition (CVPR)

作者： Kim, Junsu Ku, Yunhoe Kim, Jihyeon Cha, Junuk Baek, Seungryul UNIST Ulsan South Korea MODULABS Seoul South Korea

ISBN: (纸本)9798350365474

In the field of Class Incremental Object Detection (CIOD), creating models that can continuously learn like humans is a major challenge. Pseudo-labeling methods, although initially powerful, struggle with multi-scenario incremental learning due to their tendency to forget past knowledge. To overcome this, we introduce a new approach called vision-Language Model assisted Pseudo-Labeling (VLM-PL). This technique uses vision-Language Model (VLM) to verify the correctness of pseudo ground-truths (GTs) without requiring additional model training. VLM-PL starts by deriving pseudo GTs from a pre-trained detector. Then, we generate custom queries for each pseudo GT using carefully designed prompt templates that combine image and text features. This allows the VLM to classify the correctness through its responses. Furthermore, VLM-PL integrates refined pseudo and real GTs from upcoming training, effectively combining new and old knowledge. Extensive experiments conducted on the Pascal VOC and MS COCO datasets not only highlight VLM-PL's exceptional performance in multi-scenario but also illuminate its effectiveness in dual-scenario by achieving state-of-the-art results in both.

关键词： CIOD Class Incremental Object Detection Continual Learning Incremental Learning Object Detection Pseudo Labeling vision-Language Model

来源：评论

学校读者我要写书评

暂无评论

A Comprehensive Analysis of Factors Impacting Membership Inference

A Comprehensive Analysis of Factors Impacting Membership Inf...

引用

ieee/cvf conference on computer vision and pattern recognition (CVPR)

作者： DeAlcala, Daniel Mancera, Gonzalo Morales, Aythami Fierrez, Julian Tolosana, Ruben Ortega-Garcia, Javier Univ Autonoma Madrid Biometr & Data Pattern Analyt Lab Madrid Spain

ISBN: (纸本)9798350365474

We analyze various factors affecting the proper functioning of MIA and MINT, two research lines aimed at detecting data used for training. The difference between these lines lies in the environmental conditions, while the fundamental bases are similar for both. As evident in the literature, this detection task is far from straightforward and poses an ongoing challenge for the scientific community. Specifically, in this work, we conclude that factors such as the number of times data passes through the original network, the loss function, or dropout significantly impact detection outcomes. Therefore, it is crucial to consider them when developing these methods and during the training of any neural network, both to avoid (MIA) and to enhance (MINT) this detection. We evaluate the AdaFace facial recognition model using five databases with over 22 million images, modifying the different factors under analysis and defining a suitable protocol for their examination. State-of-the-art accuracy reaching up to 87% is achieved, surpassing existing methods.

关键词： Face recognition Fairness Membership Inference MIA MINT Realiability

来源：评论

学校读者我要写书评

暂无评论

VMCML: Video and Music Matching via Cross-Modality Lifting

VMCML: Video and Music Matching via Cross-Modality Lifting

引用

ieee/cvf conference on computer vision and pattern recognition (CVPR)

作者： Lee, Yi-Shan Tseng, Wei-Cheng Wang, Fu-En Sun, Min Natl Tsing Hua Univ Hsinchu Taiwan Univ Toronto Toronto ON Canada Vector Inst Toronto ON Canada

ISBN: (纸本)9798350365474

We propose a content-based system for matching video and background music. The system aims to address the challenges in music recommendation for new users or new music give short-form videos. To this end, we propose a cross-modal framework VMCML (Video and Music Matching via Cross-Modality Lifting) that finds a shared embedding space between video and music representations. To ensure the embedding space can be effectively shared by both representations, we leverage CosFace loss based on margin-based cosine similarity loss. Furthermore, to confirm the music is not the original sound of the video and that more than one video is matched to the same music, we follow the rule and collect videos and music from a well-known multi-media platform. That is because there are limitations of previous datasets. We establish a large-scale dataset called MSV, which provide 390 individual music and the corresponding matched 150,000 videos. We conduct extensive experiments on Youtube-8M and our MSV datasets. Our quantitative and qualitative results demonstrate the effectiveness of our proposed framework and achieve state-of-the-art video and music matching performance.

关键词： computer music

来源：评论

学校读者我要写书评

暂无评论

How Much You Ate? Food Portion Estimation on Spoons

How Much You Ate? Food Portion Estimation on Spoons

引用

ieee/cvf conference on computer vision and pattern recognition (CVPR)

作者： Sharma, Aaryam Czarnecki, Chris Chen, Yuhao Xi, Pengcheng Xu, Linlin Wong, Alexander Univ Waterloo Vis & Image Proc Lab Waterloo ON Canada Natl Res Council Canada Ottawa ON Canada

ISBN: (纸本)9798350365474

Monitoring dietary intake is a crucial aspect of promoting healthy living. In recent years, advances in computer vision technology have facilitated dietary intake monitoring through the use of images and depth cameras. However, the current state-of-the-art image-based food portion estimation algorithms assume that users take images of their meals one or two times, which can be inconvenient and fail to capture food items that are not visible from a top-down perspective, such as ingredients submerged in a stew. To address these limitations, we introduce an innovative solution that utilizes stationary user-facing cameras to track food items on utensils, not requiring any change of camera perspective after installation. The shallow depth of utensils provides a more favorable angle for capturing food items, and tracking them on the utensil's surface offers a significantly more accurate estimation of dietary intake without the need for post-meal image capture. The system is reliable for estimation of nutritional content of liquid-solid heterogeneous mixtures such as soups and stews. Through a series of experiments, we demonstrate the exceptional potential of our method as a non-invasive, user-friendly, and highly accurate dietary intake monitoring tool.

关键词： computer-vision estimation food nutrition volumetric

来源：评论

学校读者我要写书评

暂无评论

没有更多数据了...

全选清除本页清除全部题录导出标记到“检索档案”

共392页 << < 2 3 4 5 6 7 8 9 10 11 > >>

检索报告对象比较合并检索0

隐藏清空

合并搜索

回到顶部

执行限定条件

内容：

评分：

请选择保存的检索档案：

请选择收藏分类：

订阅名称：

通借通还

温馨提示：

图书名称：

借书校区：

取书校区：

手机号码：

邮箱地址：

一卡通帐号：

电话和邮箱必须正确填写，我们会与您联系确认。

联系人：

所在院系：

联系邮箱：

联系电话：

内蒙古自治区呼和浩特市赛罕区大学西街235号邮编: 010021

建议与咨询 留下您的常用邮箱和电话号码，以便我们向您反馈解决方案和替代方法

分类表

所选分类

限定检索结果

文献类型

馆藏范围

日期分布

学科分类号

主题

机构

作者

语言

请选择保存的检索档案： 新增检索档案 确定 取消

请选择收藏分类： 新增自定义分类 确定 取消

通借通还

建议与咨询留下您的常用邮箱和电话号码，以便我们向您反馈解决方案和替代方法

请选择保存的检索档案：

请选择收藏分类：