--
Best Regards...
| ||
|
|
PLDWorld 홈페이지의 유지보수를 위해, 여기저기 서핑중 발견되는 각종 자잘한 & 미쳐 정리가 되지않은 나만의 자료와 더불어 나의 "일상다반사"가 하나하나씩 저장되는 곳... 나중에 정리되는 Contents들은 그때마다 하나씩 없어질런지도... :)
--
Best Regards...
| ||
|
|
#10 Alternative computing solutions, from single cores to arrays of 'things'
There are many ways of performing computations, including single CPU or DSP processors (chips or cores), multiple processors, arrays of "things", and "great big piles of gates."
#9 The state-of-play in multi-processor and reconfigurable computing
When a conventional processor (core) cannot meet the needs of a target application, it becomes necessary to evaluate alternative solutions such as multiple cores and/or configurable cores.
#8: How to take advantage of partial reconfiguration in FPGA designs.
The capability of designs to leverage partial reconfiguration opens doors to a whole host of applications.
#7: How to invert three signals with only two NOT gates (and *no* XOR gates): Part 2.
In part two of this article, we consider a dynamic solution to our original problem (using a ring oscillator and other "stuff"); also, we learn how to implement a NOT gate using four AND gates!
#6: FPGA Architectures from 'A' to 'Z' – Part 2.
If you are new to FPGAs, there are a bewildering number of different architectures and related concepts; but fear not, because this tutorial explains all.
#5: All About FPGAs.
An industry expert examines field-programmable gate arrays (FPGAs), including current and forthcoming architectures, technologies, and software tools.
#4: How to implement a digital oscilloscope in Structured ASIC fabric.
Structured ASICs provide quicker time-to-market and lower development costs than standard ASICs, while also providing higher performance and lower unit costs than FPGAs.
#3: How to invert three signals with only two NOT gates (and *no* XOR gates): Part 1
Even for hardened logic designers, these solutions will delight and entertain; also, there's a new "Brain Boggler" to be pondered.
#2: FPGA architectures from 'A' to 'Z' – Part 1
If you are new to FPGAs, there are a bewildering number of different architectures and related concepts; but fear not, because this two-part tutorial explains all.
#1: An introduction to different rounding algorithms
The mind soon boggles at the variety and intricacies of the rounding schemes that may be used for different applications. In addition to introducing different techniques, this article provides real-world examples of the types of errors associated with the different rounding schemes applied at various stages throughout a digital filter.
Embarrassingly enough, #3, #2, and #1 are all three that were penned by yours truly . . . just give me a moment, I promised myself I wouldn't cry . . .
Clive "Max" Maxfield is the editor of Programmable Logic DesignLine. Max is the author and co-author of a number of books, including Bebop to the Boolean Boogie (An Unconventional Guide to Electronics), The Design Warrior's Guide to FPGAs (Devices, Tools, and Flows), and How Computers Do Math featuring the pedagogical and phantasmagorical virtual DIY Calculator.
Widely regarded as being an expert in all aspects of computing and electronics (at least by his mother), Max was once referred to as "an industry notable" and a "semiconductor design expert" by someone famous who wasn't prompted, coerced, or remunerated in any way. Max can be reached at max@techbites.com.
| [글로벌 IT이슈 진단]무선충전시대가 열린다 |
| [ 2007-01-31 ] |
| 요즘 휴대폰은 만능이다. 사진·e메일·동영상·TV까지 한 대면 모두 다 할 수 있다. 하지만 이런 첨단 휴대폰에도 변하지 않는 원시적인 단점이 있다. 바로 전력을 공급받는 문제, 즉 충전 부분이다. 휴대폰으로 여러 가지를 하다보니 하루라도 충전을 거른 날이면 전원이 꺼질까 노심초사한다. 언제 어디를 가도 휴대폰이 자동으로 충전될 순 없을까. 무선 전력은 상상 속에서만 가능한 일일까. ◇'5m까지 쏜다'=작년 연말 미국에선 무선 전력을 기대할 수 있는 흥미로운 논문이 발표됐다. 미국 메사추세츠공과대(MIT) 마린 솔자식 교수팀은 노트북PC, MP3플레이어 배터리를 3∼5m 떨어진 곳에서도 무선으로 충전할 수 있다고 주장했다. 이들의 얘기는 비록 컴퓨터 시뮬레이션에 근거한 것이지만 꽤 먼 거리에서도 무선 충전이 가능하다고 해 큰 관심을 모았다. 이 기술은 '공진(resonance)'을 이용했다. 공진이란 어떤 물체가 외부에서 그 고유 진동 수(초당 진동 횟수)와 같은 진동수를 가진 힘을 받으면 진폭이 증가하는 현상으로, 진동 수가 같은 두 개의 소리굽쇠를 가까이 두고 하나를 때리면 다른 소리굽쇠가 울리는 것이다. 솔자직 교수팀은 소리 대신 전기 에너지를 담은 전자기파를 공진시켜 배터리를 충전할 수 있다고 했다. 에너지 분산을 막기 위해 고안한 특수 안테나를 전원부와 노트북에 각각 설치한 후 전류를 흘리면 공진에 의해 두 개의 안테나 사이에서 에너지가 전달되고 이를 통해 충전할 수 있다고 주장했다. ◇한층 가까워진 무선 충전=솔자직 교수팀의 논문은 아직 이론적인 얘기지만 세계에선 무선 충전이 눈 앞의 현실로 나타나고 있다. 올 초 라스베이거스에서 열린 가전전시회 'CES 2007'에선 세계 여러 기업들이 이를 증명해 보였다. 애리조나에 있는 와일드차지와 미시간 풀톤 이노베이션은 어댑터가 없어도 휴대 가전들을 충전할 수 있는 제품을 선보였다. 와일드차지는 휴대폰·MP3플레이어·디지털 카메라·노트북 등 다양한 종류의 휴대기기들을 마우스패드와 같은 장치 위에 올려 놓으면 동시에 충전할 수 있는 '와일드차저(WildCharger)'를, 풀톤 이노베이션은 유도전류를 응용한 무선 충전 기술 '이커플드(eCoupled)'를 발표했다. 이커플드의 원리는 구체적으로 공개되지 않았지만 전동칫솔의 충전 방식과 같은 것으로 전해졌다. 이 밖에 파워캐스트는 900㎒ 대역 RF통신을 통해 밀리와트(mW) 단위의 작은 전력을 최대 1m까지 전달하는 기술을 선보였다. ◇꿈은 이뤄지나=무선 충전, 즉 무선으로 전력을 전송하기 위한 시도가 최근의 일은 아니다. 에디슨을 뛰어 넘는 천재성에도 불구하고 제대로 평가를 받지 못한 19세기 물리학자, 니콜라 테슬라는 대형 안테나를 통해 전기를 무선으로 전달하는 실험을 한 바 있으며, 영국의 스플래쉬파워는 2002년 전자기파를 이용해 휴대폰·MP3플레이어 등을 충전하는 패드를 일찍이 고안하기도 했다. 하지만 원가·발열·효율 등의 문제로 이 같은 무선 충전 기술이 그동안 상용화되진 못했다. 그러나 상황이 달라지는 것처럼 보인다. 가전 업체들이 신기술들을 적극 채택하고 나섰기 때문이다. 파워캐스트는 필립스와 제휴를 맺고 연내 필립스 가전 제품에 자사의 충전 기술을 탑재할 것으로 알려졌으며 풀톤 이노베이션은 세계적인 자동차 전장 업체인 비스테온과 함께 휴대기기를 충전할 수 있는 차량용 무선충전 장치를 올 여름 내놓을 계획이다. 풀톤은 또 모토로라·모빌리티 일렉트로닉스와 제휴를 맺고 자사의 무선 충전 기술을 산업 표준으로 만들려는 움직임도 보이고 있다. 새롭게 등장한 무선 충전 기술들이 얼마나 효용성이 있고 또 시장에서 성공할 수 있을 지 현재로선 예측하기 어렵지만 모바일 시대를 맞아 충전 기술에도 변화가 시작됐음은 분명해 보인다. ◆'내 휴대폰도 무선 충전한다?' '내 휴대폰도 무선으로 충전하는 방법이 있다?' 물론 있다. 단 '011'과 '삼성' 휴대폰 사용자야 한다. 먼저 '5, 8, 0, 9, 5, 4, 0'을 순서대로 누른 후 '*'와 '4, 5, 6, 8, 0'을 추가로 입력한다. 또 이어서 '7, 0, 1, 1, 4'와 휴대폰의 '확인' 버튼을 누르고 마지막으로 '0, 2'를 입력하면 휴대폰 배터리 눈금에 불이 들어온 것(완충)을 볼 수 있다. 그러나 이는 실제로 배터리를 충전시키는 것이 아니라 충전된 것처럼 보이게 하는 방법이다. 제조사가 휴대폰을 출하하기 전 품질 검사를 위해 테스트 모드로 들어가 확인하는 것이다. 삼성전자는 "테스트를 위해 전압을 낮추는데 기준 전압이 낮아지면서 순간적으로 배터리 눈금이 다시 올라가는 것일 뿐"이라고 설명했다. 인터넷에서는 이 번호들이 '배터리를 무선 충전하는 법'으로 소개되고 있지만 잘못 다루면 휴대폰을 고장낼 수 있다. ◆국내 현황은 특허청 산하 특허정보원 사이트에서 '무선 충전'으로 관련 특허를 찾아보니 총 10건이 검색됐다. 이 중 최종 특허로 등록된 것은 6건뿐이었고 나머지는 현재 공개 중이거나 포기 또는 거절된 상태였다. 무선 충전에 관한 국내 기술 개발은 표면적으론 그리 활발해 보이지 않았다. 하지만 국내에도 무선 충전 기술을 집중적으로 연구하는 회사가 있다. 지난 2001년 안동대학교 교수 벤처기업으로 설립된 '제이씨프로텍'이다. 현 안동대 전자공학 교수이자 JC프로텍 대표인 이형주 사장이 개발한 것은 미국 와일드차지, 영국 스플래시파워와 유사한 '무선 충전 패드'다. JC프로텍의 제품은 코일로 감싸진 송신부 역할의 패드와 같은 코일로 된 수신 모듈로 구성됐다. 송신부에 전원을 넣으면 이곳에서 나오는 방사 전파의 자기장을 수신 코일부에서 받아 들여 패러데이 전자기 유도 법칙에 의해 전기로 변환, 배터리가 충전되는 방식이다. 이런 원리를 통해 패드 위에 휴대폰이나 MP3플레이어 등을 놓기만 하면 자동 충전이 된다. 특히 최대 네대까지 휴대기기들을 동시에 충전할 수 있어 여러 개의 어댑터를 쓸 일이 없어진다. 이 사장은 "자기장의 전달 효율이 낮아 해외업체들도 이를 구현하는 것이 쉽지 않았는데 우리는 특수 설계한 코일 구조로 80∼90% 효율을 확보했다"면서 "충전 시간도 유선과 동일하고 사용 시간도 유선과 똑같다"고 강조했다. 무선 충전 상용화의 또 다른 관건이던 가격 문제도 해소해 패드 1개와 수신 모듈이 내장된 외장형 배터리 2개를 총 2만∼3만원에 시판할 수 있을 것으로 예상한다고 이 사장은 말했다. 현재 JC프로텍은 아이팟용으로 제품을 출시하기 위해 미국 모 기업과 제휴를 논의하고 있으며 6개월 후 상품을 출시할 예정이라고 덧붙였다. 그는 자기장에 따른 인체 영향에 대해서는 "FCC 및 CE, 스웨덴의 전자파 규제를 모두 만족시켜 문제가 없다"고 강조하며 "50cm 정도 떨어져 있어도 충전이 가능한 기술도 연구 중에 있다"고 전했다. 윤건일기자@전자신문, benyun@ |
| Copyrightⓒ 2000-2005 ELECTRONIC TIMES INTERNET CO., LTD. All Rights Reserved. |
-- -- / Chang-woo YANG / --------------------------- |
| Original Page |
Page URL: http://www.altera.com/support/ip/processors/nios2/ide/ips-nios2-ide-tutorial.html
The Nios II Integrated Development Environment (IDE) provides a searchable, indexed online help system that covers all IDE-related topics. The online help also includes tutorials and FAQs. These help topics are accessible once you install the Nios II embedded processor development tools. The online help is also available in PDF format on the Nios II Literature page.
The Nios II IDE Quick-Start Tutorial:
To start the Nios II IDE Quick-Start Tutorial:
The Nios II Software Development Tutorial:
To access the tutorial:
In addition, you can use the following related tutorial on using software components with Nios II processors:
The IDE includes FAQs and troubleshooting tips that provide information to assist you in developing Nios II software.
To access the FAQs and troubleshooting tips in the IDE:
Copyright © 1995 - 2007 Altera Corporation, 101 Innovation Drive, San Jose, California 95134, USA.
뉴스 및 동향 | |
수익성 높은 새로운 디자인의 아이팟 나노 모델 By David Carey Apple사는 2세대 아이팟 나노를 내놓으면서 자사의 인기 있는 플래시 기반 뮤직 플레이어의 디자인을 바꿨다. 전체 아이팟 제품 라인의 중간에 위치하고 있는 2세대 나노는 2기가바이트, 4기가바이트 및 8기가바이트 용량의 모델이 제공되는데, 이들은 각각 500곡, 1,000곡 및 2,000곡 정도의 노래를 저장할 수 있는 용량에 해당한다. 2세대 제품에서 달라진 것은 무엇일까? 가장 눈에 띄는 요소들로부터 시작하자면, Apple사는 케이스 디자인을 재구성하여 스테인리스제 통과 플라스틱제 프론트 케이스의 2부분으로 되어 있던 방식을 버리고 양극화 알루미늄 쉘 하나로 된 형태를 채택했다. 이 새로운 알루미늄 몸체 속에 전자 장치들이 빽빽하게 들어차 있다. 케이스 끝의 캡 부분들을 떼어낸 뒤 (아주 작은) 나사 몇 개를 풀어내면 핵심 전자 장치들을 케이스 통의 한쪽 끝으로부터 끄집어 낼 수 있다. 케이스의 대략적인 구조는 지금은 단종된 Apple사의 아이팟 미니로부터 빌려온 것이다. 마이크로드라이브 기반의 아이팟 미니는 상대적으로 상당히 큰 케이스를 필요로 했으며 전자 패킹도 2세대 나노 모델에서 볼 수 있는 집적도 수준에 미치지 못했지만, 기술 자체는 유사하다. 보호용의 끝 부분 캡을 제외하고는 미니의 경우와 마찬가지로, LCD(이 경우에는 176x132 픽셀의 TFT 디스플레이)를 위한 투명한 아크릴 윈도와 스크롤휠 방식의 어셈블리가 케이스 표면의 연속성을 깨는 유일한 부분들이다. PortalPlayer와의 차이점 Apple사에서는 기계 재설계와 함께 이 제품의 공급 체인도 크게 변경했지만, 선택된 설계 요소들은 상당 부분 그대로 놔두었다. PortalPlayer를 코어 미디어 프로세서로서 사용하지 않게 된 것이 아마도 가장 큰 변화일 것이다. Apple 라벨이 붙은 ASIC인 S5L870-B05는 삼성의 제품으로서 모든 오디오 및 스틸 이미지의 디코딩을 맡고 있다. 나노 모델의 CPU에 Apple사의 독점 마킹이 찍혀 있긴 하지만, 라벨을 볼 때 이 삼성 칩 내에 ARM 코어가 탑재되어 있음을 알 수 있다. 6밀리 x 6밀리 이하 크기의 이 다이는 PortalPlayer를 기반으로 하는 나노 모델의 이전 제품과 유사한 언더필 방식의 BGA 패키지로 되어 있다. 그러나 SST사의 별도 NAND 컨트롤러 구성요소를 갖추고 있던 1세대 디자인과는 달리, 이 삼성 CPU는 NAND 인터페이스를 직접 통합시킴으로써 비용과 복잡성을 줄인 것같다. 이 삼성 프로세서는 1메가바이트의 SST 플래시를 코드 일부나 전체를 저장하는 데 사용하고 있으며, Qimonda 32메가바이트 SDRAM이 시스템의 동작 메모리를 제공하는 것은 물론 어쩌면 NAND 플래시로부터 꺼내온 노래 데이터를 위한 버퍼 역할까지도 하고 있는 것같다. 다른 2세대 나노 모델들에 대한 검사 자료를 토대로, 삼성은 32메가바이트 SDRAM의 공급업체로 알려져 있다. 첫번째 나노 모델로부터 이어져 내려온 디자인 요소들 가운데는 Cypress사의 스크롤 휠 컨트롤러(CY8C21001A)와 전력관리용으로 수정된 NXP 칩(PCF50635)이 있다. 역시 오리지널 나노 모델로부터 수정된 Wolfson사의 부품(WM8750S)은 오디오 코덱 및 헤드폰 앰프용으로 선택되어 있다. 후자의 두 부분들은 삼성의 미디어 CPU와 같이 Apple사 고유 라벨이 찍혀 있는데, 이는 아마도 공급체인을 식별하기 힘들게 하려는 시도인 것같다. 보다 명백하게 마킹 되어 있는 다른 전력관리 부품들 가운데는 National Semiconductor사의 DC/DC 컨버터(LM34910B)와 Linear Technology사의 LTC4066(이것은 나노 모델의 USB 인터페이스를 통해 2.2밀리 x 33밀리 x 55밀리 크기의 웨이퍼 형 폴리머 배터리를 충전한다)이 있다. 배터리 자체는 표준 싱글셀 3.7V에서 330밀리암페어시 정도의 용량을 공급하는 것으로 추정된다. 이것은 총 사용가능 에너지 1.2와트시에 해당한다. 비용 면에서는 새로운 삼성의 프로세서(아마도 Apple사가 5 달러 정도에 소싱했을 것으로 보인다)보다 내부 NAND 플래시 메모리에 더 많은 비용이 든다. 최저 밀도의 2기가바이트급 나노 모델에서 조차도 하이닉스의 HY27UV08AG5M 4칩 스택형 NAND(단일 TSSOP 패키지에 들어 있는)는 20 달러 정도의 비용을 차지하고 있는 것같다. 물론 이는 Apple사에 대한 가격 할인과 계약구매 타이밍에 크게 의존하지만 말이다. SDRAM의 경우와 비슷하게, NAND 플래시에서도 멀티소싱이 중요하다. Toshiba와 삼성은 둘 다 이 새로운 나노 계열을 위한 NAND 플래시의 추가 공급 업체들로 알려져 있다. 이 제품은 TSSOP 패키지의 스택들을 사용하여 8기가바이트의 보다 큰 메모리를 갖추고 있는데, 이 패키지들 각각은 내부에 독자적인 멀티다이 스택들을 포함하고 있다. 시각적 증거와 Staktek사가 공개적으로 제공한 정보를 토대로 해볼 때, Toshiba사는 Staktek사의 TSSOP 패키지 스태킹 기술을 사용하고 있는 것같다. 이 모든 것들을 고려해보면, 2기기바이트 용량의 이 2세대 나노 모델은 액세서리(이어폰, USB 케이블 및 도킹 어댑터)를 포함하여 65달러 범위의 직접생산 및 소재 비용을 가질 것으로 추정된다. 보다 고밀도의 NAND 스택 가격을 다소 높이 잡는다면 4기가바이트 및 8기가바이트 버전의 소재 및 생산비는 각각 87달러 및 132달러 범위가 될 것으로 추정할 수 있다. 세 가지 모델들(2기가바이트, 4기가바이트 및 8기가바이트)의 소매가는 150달러, 200달러 및 250달러로서, 총 마진은 좋아보인다. 이는 로엔드의 경우 56퍼센트에서 하이엔드의 경우 47퍼센트에 이른다. 물론 제품 개발, 마케팅, 출하 및 모든 소프트웨어 라이선스와 관련된 다른 간접 비용들은 이 수치에 포함되어 있지 않지만, 어떠한 경우이든 여전히 상황은 상당히 긍정적이다. 보다 고밀도 NAND의 가격이 다소 떨어진다고 해도, 탑엔드 모델의 마진은 보다 신속하게 개선될 것이다. 최신 나노 모델이 이전 모델에 필적하는 인기를 누릴지는 아직 두고 볼 일이다. 소비자들은 아이팟에 다소 싫증을 느낄지도 모르고, 이 제품 라인의 빠른 진화는 구매자들이 잦은 제품 변화 및 업그레이드를 따라올 수 있는 능력을 소진시켜 버릴지도 모른다. 하지만 시장에 대한 그 같은 질문에는 오직 시간만이 대답해 줄 수 있다. 아직까지는 Apple사가 멋진 디자인에다가 수익성도 높아 보이는 또 다른 아이팟 모델들을 가지고 여전히 오디오 플레이어 시장을 지배하고 있다.
Apple사에서는 기계 재설계와 함께 이 제품의 공급 체인도 크게 변경했지만, 선택된 설계 요소들은 상당 부분 그대로 놔두었다.
| |
Understanding the architecture
When evaluating a new FPGA architecture, it is important to understand the hardware features and the tradeoffs that can be made in the architecture. Datasheets, user guides, and technical papers on the architectural features should be thoroughly reviewed before moving forward with a design.
The first thing to learn about any FPGA is what makes up the basic fabric of logic. For example, each of the configurable logic blocks (CLBs) in a Xilinx Virtex-5 FPGA contains two slices; and each slice contains four 6-input look-up tables (LUTs), four registers, and dedicated carry logic. For maximum utilization of each slice, it is important to take into consideration the width of the LUTs, the connectivity between the basic elements, and any shared resources.
Many FPGA architectures also contain hard IP blocks, such as embedded memory and blocks used for DSP functions. If a hard-IP block continuously shows up as the source or destination of your critical path, there are a couple of things that can be analyzed to improve the performance. First, check to see if the design is making the most of the block's features and that the synthesis tool is inferring the features you expected from your RTL code. Use the dedicated pipeline registers inside the blocks to reduce the setup and clock-to-out timing. Evaluate the tradeoff between using dedicated blocks versus implementing the same function in slices to allow for placement flexibility. This can especially be important when using a high percentage of hard-IP blocks.
The clocking resources that are utilized in a design can also affect a design's performance. For example, Xilinx Virtex-5 FPGAs have I/O, regional, and global clocking resources. These devices are divided into clock regions which at most, can contain 4 regional clocks and 10 global clocks. During design planning, it is important to analyze how many clock regions are going to be used as well as specific clocks within a clock region. Placing your I/Os so that their interface logic does not require all the clock resources in a clock region gives the implementation tools greater placement flexibility.
Define timing requirements
Synthesis and implementation tools are driven by the performance goals that a user specifies with timing constraints. It is important to constrain all internal clock domains, input and output (I/O) paths, multi-cycle paths, and false paths. Define realistic timing constraints in synthesis order to prevent excessive replication.
In your synthesis report, check for any replicated registers and ensure that timing constraints that might apply to the original register also cover the replicated registers for implementation. When writing timing constraints for implementation, group the maximum number of paths with the same timing requirement first before generating a specific timing constraint. By consolidating constraints, implementation runtime and memory usage can be minimized.
Example of non-consolidated constraints (Xilinx constraint syntax)
TIMESPEC "TS_firsttimespec" = FROM "flopa" TO "flopb" 10ns;
TIMESPEC "TS_secondtimespec" = FROM "flopc" TO "flopb" 10ns;
TIMESPEC "TS_thirdtimespec" = FROM "flopd" TO "flopb" 10ns;
Consolidation of constraints using grouping
INST "flopa" TNM = "flopgroup";
INST "flopc" TNM = "flopgroup";
INST "flopd" TNM = "flopgroup";
TIMESPEC "TS_consolidated" = FROM "flopgroup" TO "flopb" 10ns;
Driving synthesis
For a synthesis tool to create a high-performance circuit, the tool needs to be properly driven by the designer. The first thing a designer needs to consider is proper coding techniques to ensure that inference of behavioral RTL made by the synthesis tool leads to the maximum usage of the architectural features. For example, Xilinx ISE Project Navigator's language templates – available in both Verilog and VHDL – are a great place to get coding examples.
Next, make sure that the synthesis tool has a complete picture of the design. If a design contains IP netlists or any other lower level black-boxed netlists, these netlists should be included in the synthesis project. Although the synthesis tool won't optimize any logic within the netlist, it will have a better understanding of how to optimize the HDL that interfaces to these lower level netlists.
The tool also needs to understand the performance goals of a design using the timing constraints supplied by the designer. If there are critical paths in the implementation that are not seen as critical in synthesis, try Synplicity Synplify PRO's –route constraint to force synthesis to focus on that path. Finally, there are a variety of tool settings in synthesis that should be explored. Refer to Fig 1 for suggested tool settings for Synplify PRO.
* For a complete listing of attributes and their functionality, please see the synthesis tool's documentation.
Although timing performance might be enhanced, options that do lead to the replication of logic such as retiming in Synplify PRO can impact area. If the design is affected by high-fanout nets and you want the synthesis tool to reduce that fanout, use fanout attributes specifically on that specific net, versus globally specifying a maximum fanout limit. If hierarchical boundaries are maintained, a designer should make sure that ports are registered at the hierarchical boundaries. If critical paths cross over these hierarchical boundaries, certain optimizations will not be allowed by the synthesis tool. This can lead both to lower performance and higher area utilization. Before moving on to implementation, it is always important to review the warnings in the synthesis report. It is also beneficial to check the RTL schematic view for how the synthesis tool is interpreting the HDL and the technology schematic to understand how the HDL is mapping to the specific FPGA architecture.
Choosing implementation options
Having obtained an acceptable timing estimate from the synthesis tool, use the implementation tools to determine the true performance of the design. The implementation options that can be used are unique for each design depending on the performance goals of the design, the synthesis flow used, and its overall structure. Once the majority of the functionality is defined in the design's HDL and the effort is focused on timing closure, it is beneficial to run a series of different implementations with different sets of options to determine which is the best combination for the design.
ISE Xplorer is an example of a tool that will allow a designer to determine which options work best. ISE Xplorer has been tuned for each Xilinx FPGA architecture to try the best set of combinations. Although initial runtime can be longer because multiple implementations need to be run, once the design has the right set of options, it will likely reduce the number of design iterations to achieve timing closure.
Physical synthesis options in implementation can be used to re-optimize and pack logic based on knowledge of the critical paths of a design, leading to better placement and routing. Note that physical synthesis can lead to increased area due to replication of logic. Like synthesis, if hierarchy is maintained on a design but the critical path crosses those hierarchical boundaries, physical synthesis will not be able to optimize that path and potentially, inefficient packing will occur. To evaluate whether keeping hierarchy is affecting the performance of the design, turn off hierarchy preservation with an attribute or option during implementation. If it does prove to have an impact, reconsider how the hierarchical boundaries are defined.
Evaluating critical paths
By understanding the characteristics of the critical path, a designer can make better decisions on what to do for the next design iteration. A data path is comprised of both logic and interconnect delay. Individual component delays that make up logic delay are fixed. Logic delay can only be reduced if the number of logic levels are reduced or the structure of the logic is changed. By comparison, interconnect delay is much more variable and is dependent on the placement of the logic, routing congestion, and the competition between nets for the fastest routing resources. Before routing the design a quick timing analysis after placement is recommended. Although this timing report will only have estimates for the routing delays, it will give an idea of the critical paths the implementation tools are working on. If the critical paths have a high number of logic levels, designers may want to work on improving the logic levels versus running it through PAR. When the design has an excessive amount of logic levels that lead to many routing interconnects:
In the case where there are few logic levels but the certain data paths are not meeting the performance requirement:
Conclusion
Today's FPGAs have a variety of high performance features. In order to take full advantage of these features, a few things need to be considered. The more that can be done upfront with good coding styles, timing constraints definition, and resource planning, the easier it will be for the downstream tools to achieve timing requirements. It is also equally important to know what to do next when design requirements are not met in first iteration.
Michelle Fernandez is a technical marketing engineer in the Software Product Marketing Group at Xilinx. Based on the analysis of customer designs, Michelle provides recommendations aimed at improving FPGA design performance and ease of use to the development teams at Xilinx. Michelle joined Xilinx in 1999 and has held a variety of positions in customer support and field applications engineering. She holds a bachelor's of science degree in electrical engineering from University of California at Davis. Michelle can be contacted at: michelle.fernandez@xilinx.com.